Tencent claims its released EVIE visual-document retrieval models achieve 66.75 nDCG@10 on ViDoRe V3 using 4096-dimensional per-token embeddings that preserve layout, charts, and tables, potentially improving retrieval for visually structured document RAG.
state: seedheat: lowuncertainty: highknownscott: lowvisual-document-retrieval multimodal-rag open-modelsTencent
What is this?
Tencent has a visual-document retrieval model page on Hugging Face for EVIE-Preview-4.5B, which claims top ViDoRe rankings and native 128-dimensional token vectors. ViDoRe V3 evaluates multimodal document RAG, including interpretation of tables, charts, and images, cross-document synthesis, and source grounding; its retrieval metric includes nDCG@10. The supplied search summary repeats the case’s 66.75 score and 4096-dimensional embedding claim, but the underlying snippets do not substantiate that score or an EVIE-8B release, and the Tencent page instead specifies 128-dimensional vectors for the preview model. The claimed release details and layout-preservation benefits therefore remain unverified in this material.
Why it matters to Scott
The radar already tracks this development in radar:tencent-evie-128d-visual-retrieval; the supplied grounding does not verify the new 66.75 score, 4096-dimensional vectors, or EVIE-8B release. Visual retrieval touches Scott’s “Text Is the Model’s Home Turf” distinction—retain pixels when layout carries meaning—but this case supplies neither a validated comparison that changes that boundary nor a demonstrated improvement to his workflows.
ip:concept.text-is-the-models-home-turfradar:tencent-evie-128d-visual-retrieval
queries asked of Scott's wikis
- visual document RAG versus OCR text extraction
- multivector retrieval token embeddings index cost
- chart table layout preservation knowledge ingestion
- multimodal retrieval evaluation source grounding
- self-hosted embedding models retrieval infrastructure
Measured heat
now 0 pts/hpeak 0 pts/hcomments 0/hpeers p0momentum: steady1 platformsage 822h
points/hour across evidence · reading as of 2026-10-12 02:59:37.977291+11:00 · deterministic, not a model opinion
How the heat travelled
pace: p64 vs 519 stories at the 720h mark (now 822h old) — ahead of hunterbench-live-pentesting-benchmark (1.0x), behind google-finland-15b-ai-investment (1.0x)
Evidence (1) — ⭐ canonical anchor
Interpretation history
2026-09-09T15:36:21Z
The latest look adds no substantive evidence beyond the already-recorded configuration caveat; this remains an unvalidated high-capacity extension of the tracked EVIE preview, not a demonstrated improvement for document RAG. Further engagement does not establish configuration-specific accuracy or practical index costs.
2026-09-07T15:25:32Z
A commenter citing the model card introduces a concrete configuration caveat: the reported 3.81 GiB setup scores 59.58 on V3, rather than 66.02, suggesting compact-index performance must be separated from headline accuracy. This is an unverified reading, not an independent evaluation or a direct rebuttal of the 66.75 claim; practical retrieval-versus-storage gains remain unsettled.
2026-09-07T10:28:56Z
No substantive evidence has been added: the reported high-capacity EVIE variant remains an unverified extension of the already tracked preview release, not independent corroboration. The preview's 128-dimensional specification does not disprove a separate 4096-dimensional variant, but neither the variant nor its practical retrieval-versus-storage trade-off is established here.
2026-09-07T10:25:34Z
grounded: known/low — The radar already tracks this development in radar:tencent-evie-128d-visual-retrieval; the supplied grounding does not verify the new 66.75 score, 4096-dimensio
2026-09-07T10:22:55Z
case created — A linked Tencent model artifact and a specific retrieval score establish a bounded release worth tracking, although practical quality and embedding-storage costs remain unvalidated.
Decision trace
- 10-05 20:23review_dormant28 days without material information; scheduled checks stopped
- 09-10 01:36repriceThe latest look adds no substantive evidence beyond the already-recorded configuration caveat; this remains an unvalidated high-capacity extension of the tracked EVIE preview, not a demonstrated impro
- 09-10 01:36alert_silentThere is no new release, access change, evaluation, or implementation in this delta. The unresolved configuration trade-off can wait for a briefing rather than consume Scott's attention today.
- 09-10 01:36alert_routeThere is no new release, access change, evaluation, or implementation in this delta. The unresolved configuration trade-off can wait for a briefing rather than consume Scott's attention today.
- 09-09 03:21sensor_dirtyengagement_update
- 09-08 20:21sensor_dirtyengagement_update
- 09-08 18:21sensor_dirtyengagement_update
- 09-08 14:21sensor_dirtyengagement_update
- 09-08 11:21sensor_dirtyengagement_update
- 09-08 10:21sensor_dirtyengagement_update
- 09-08 06:21sensor_dirtyengagement_update
- 09-08 03:22sensor_dirtyengagement_update
- 09-08 01:25repriceA commenter citing the model card introduces a concrete configuration caveat: the reported 3.81 GiB setup scores 59.58 on V3, rather than 66.02, suggesting compact-index performance must be separated
- 09-08 01:25alert_silentThe new configuration caveat is useful for a future evaluation but establishes neither a validated deployment trade-off nor a change requiring Scott's attention today. It can wait for the next br
- 09-08 01:25alert_routeThe new configuration caveat is useful for a future evaluation but establishes neither a validated deployment trade-off nor a change requiring Scott's attention today. It can wait for the next br
- 09-08 01:21sensor_dirtycomment_update
- 09-08 00:21sensor_dirtyengagement_update
- 09-07 22:21sensor_dirtyengagement_update
- 09-07 20:28repriceNo substantive evidence has been added: the reported high-capacity EVIE variant remains an unverified extension of the already tracked preview release, not independent corroboration. The preview'
- 09-07 20:28alert_silentThe new observation adds no release confirmation, implementation, or benchmark validation. Existing coverage already captures the visual-retrieval development, and this unresolved variant claim can wa
- 09-07 20:28alert_routeThe new observation adds no release confirmation, implementation, or benchmark validation. Existing coverage already captures the visual-retrieval development, and this unresolved variant claim can wa
- 09-07 20:28alert_silentThe post links Tencent model repositories and reports concrete EVIE-8B/4.5B retrieval scores and embedding configurations, making it useful briefing material. However, the supplied evidence is a secon
- 09-07 20:28alert_routeThe post links Tencent model repositories and reports concrete EVIE-8B/4.5B retrieval scores and embedding configurations, making it useful briefing material. However, the supplied evidence is a secon
- 09-07 20:25groundThe radar already tracks this development in radar:tencent-evie-128d-visual-retrieval; the supplied grounding does not verify the new 66.75 score, 4096-dimensional vectors, or EVIE-8B release. Visual
- 09-07 20:22createA linked Tencent model artifact and a specific retrieval score establish a bounded release worth tracking, although practical quality and embedding-storage costs remain unvalidated.