Perplexity AI Releases pplx-embed-v2-late: A 0.6B Edge Model and a 9B Model Scoring 92.4% on MADQA
Quick Answer
Perplexity AI has launched the pplx-embed-v2-late model series, featuring a 0.6B edge model and a 9B model that achieves a 92.4% score on the MADQA benchmark.
Quick Take
Both models are MIT-licensed, suitable for self-hosting, and designed for different use cases in AI applications.
Key Points
- The 0.6B model is optimized for edge devices.
- The 9B model is intended for high-quality index creation.
- Best performance recorded is 92.4% on MADQA.
- Lowest score achieved is 61.2% on ViDoRe v3 Markdown.
- Both models are available under MIT licensing.
Article Excerpt
From source RSS / original summaryPerplexity's pplx-embed-v2-late comes in 2 sizes: a 0. 6B model built to run on edge devices, and a 9B model for building high-quality indexes. Its best score is 92. 4% on MADQA, and its weakest is 61. 2% on ViDoRe v3 Markdown. Both are MIT-licensed and ready to self-host. The post Perplexity AI Releases pplx-embed-v2-late: A 0. 6B Edge Model and a 9B Model Scoring 92. 4% on MADQA appeared first on MarkTechPost.
Want this in your inbox every morning?
Daily brief at your local 8am — bilingual EN/中文, free.
More from MarkTechPost
See more →Meet Flash-KMeans: An IO-Aware, Exact K-Means That Runs Over 200× Faster Than FAISS on GPUs
Flash-KMeans is an open-source, IO-aware k-means implementation that operates over 200× faster than FAISS on NVIDIA H200 GPUs. It achieves 17.9× end-to-end and 33× speedup over cuML by optimizing distance calculations and updating mechanisms without approximating results. This advancement significantly enhances performance for data scientists and machine learning practitioners.