Articles & Guides

Claude API relay guides, detection insights and hands-on LLM API benchmarks

791 articles

CCTest · Blog
Selective Risk Control for Document Extraction Has a Validity Problem
Evaluation & Benchmarks
cctest.ai

Selective Risk Control for Document Extraction Has a Validity Problem

A study of real receipt fields finds that ordinary confidence-thresholding can miss its risk target because document fields are clustered, scores leak into threshold fitting, and discrete scores create unstable thresholds. It proposes a validity ladder and identifies when conditioning improves coverage rather than merely fragmenting the calibration set.

Read more
CCTest · Blog
SPARGen unifies 3D reconstruction, dense correspondence, and spatial reasoning in one multimodal generator
Multimodal
cctest.ai
Multimodal

SPARGen unifies 3D reconstruction, dense correspondence, and spatial reasoning in one multimodal generator

SPARGen reframes several spatial intelligence tasks as instruction-conditioned generation within a native multimodal model. Its main contribution is a unified interface where geometric, correspondence, and reasoning supervision can shape shared representations.

Read more