⚡ AI NEWS WIRE · Asif Razzaq · September 6, 2026
We look at NeoMME, a family of 260M and 800M bidirectional encoders from H Company. Unlike ColPali-style retrievers, it processes multilingual text tokens and raw 32×32 image patches in a single Transformer, with no pretrained vision tower and no causal decoder. We cover the masked discrete-diffusion pretraining objective, the dual dense and late-interaction retrieval heads, and the ViDoRe v3 results where the 260M model reaches 0.523 nDCG@10. We also break…
Source: MarkTechPost — read the full story →
📌 We track AI news like this daily on the GPTAIClips YouTube channel — new tools installed live, with free playbooks at gptaiclips.com/guides.