Articles & Guides

Claude API relay guides, detection insights and hands-on LLM API benchmarks

795 articles

CCTest · Blog
Xiaomi-Robotics-1 Brings Scaling Laws to Robot VLA Models with 100K+ Hours of Real Trajectories
Robotics & Physical AI
cctest.ai

Xiaomi-Robotics-1 Brings Scaling Laws to Robot VLA Models with 100K+ Hours of Real Trajectories

Xiaomi-Robotics-1 explores whether robotics can benefit from the same scaling dynamics that transformed language and vision models. By pre-training on more than 100,000 hours of real manipulation trajectories, the model shows consistent gains as data and model size increase.

Read more
CCTest · Blog
Robot-Centric Pointmaps Help VLA Models See From the Robot’s Frame
Robotics & Physical AI
cctest.ai

Robot-Centric Pointmaps Help VLA Models See From the Robot’s Frame

A KAIST AI paper introduces robot-centric pointmaps to reduce the mismatch between camera observations and robot-frame actions in VLA models. The method encodes 3D scene coordinates in the robot frame while keeping the image-like grid structure used by existing 2D vision backbones.

Read more
CCTest · Blog
OPD² Reframes On-Policy Distillation Around What Reasoning Tuning Adds
Reinforcement Learning
cctest.ai

OPD² Reframes On-Policy Distillation Around What Reasoning Tuning Adds

On-Policy Delta Distillation, or OPD², replaces direct imitation of a teacher’s full output distribution with a signal based on the difference between a reasoning-tuned teacher and its base model. The paper reports consistent gains across math, science, and code-reasoning benchmarks.

Read more
CCTest · Blog
S1-Omni Unifies Scientific Multimodal Reasoning Across Molecules, Proteins, Materials, and Images
AI for Science
cctest.ai
AI for Science

S1-Omni Unifies Scientific Multimodal Reasoning Across Molecules, Proteins, Materials, and Images

S1-Omni aims to reduce fragmentation in AI for Science by bringing diverse scientific data types into one reasoning model. It supports more than 200 scientific tasks and releases model weights, inference code, and a 10K-sample corpus subset.

Read more