Liquid AI Releases Open-Weight d1-3B and d1-omni-600M: Multimodal Decision Models With Zero Output Tokens
Quick Answer
Liquid AI has launched two open-weight multimodal decision models, d1-3B and d1-omni-600M, which process text and images or audio without generating output tokens.
Quick Take
These models aim for real-time decision-making, enhancing efficiency in various applications.
Key Points
- d1-3B processes text and images, while d1-omni-600M handles text with images or audio.
- Both models return calibrated answers in a single forward pass.
- Zero output tokens enhance efficiency for real-time decision-making.
- These models are part of the d1 decision model family from Liquid AI.
- The release targets applications requiring quick and accurate multimodal responses.
Article Excerpt
From source RSS / original summaryLiquid AI has released Open d1, two open-weight multimodal models in its d1 decision model family. d1-3B reads text and images. d1-omni-600M reads text with an image, or text with audio. Neither model writes text. Each returns calibrated, typed answers in one forward pass with zero output tokens. The target is real-time decisions on the […] The post Liquid AI Releases Open-Weight d1-3B and d1-omni-600M: Multimodal Decision Models With Zero Output Tokens appeared first on MarkTechPost.
Want this in your inbox every morning?
Daily brief at your local 8am — bilingual EN/中文, free.
More from MarkTechPost
See more →Meet Flash-KMeans: An IO-Aware, Exact K-Means That Runs Over 200× Faster Than FAISS on GPUs
Flash-KMeans is an open-source, IO-aware k-means implementation that operates over 200× faster than FAISS on NVIDIA H200 GPUs. It achieves 17.9× end-to-end and 33× speedup over cuML by optimizing distance calculations and updating mechanisms without approximating results. This advancement significantly enhances performance for data scientists and machine learning practitioners.