Reducto Releases r-1: A Single Pass Document Parsing Model That Cuts Errors 20% at 1 Cent Per Page
Source: MarkTechPost Last week, Reducto announced r-1. It is the first model in a new parsing family built...
OpenBMB Releases MiniCPM5-2B: A 2.52B Dense Model Averaging 53.9 Across 34 Benchmarks and Built to Run On Device
Source: MarkTechPost OpenBMB has released MiniCPM5-2B, the second checkpoint in the MiniCPM5 series and the follow-up to MiniCPM5-1B....
Axis Robotics Releases AXIS: A Browser-Based Data Engine With 207 Robot Manipulation Tasks and 50,129 Trajectories
Source: MarkTechPost Robot manipulation datasets have grown far slower than the models trained on them, mostly because collection...
IFM Releases K2 Horizon: Six Apache 2.0 Models From 0.9B to 375B
Source: MarkTechPost Most open model launches release one checkpoint and a benchmark table. The Institute of Foundation Models...
H Company Releases NeoMME: A Family of 260M and 800M Single-Tower Multimodal Encoders That Drop the Vision Tower and Causal Decoder
Source: MarkTechPost Most visual document retrievers in production today are hand-me-downs. ColPali and the models that followed it...
Meta FAIR Introduces AI Research Preference Models (RPMs): Ranking ML Experiments Before Spending GPU Hours
Source: MarkTechPost AI research agents can already propose, implement and score their own machine learning experiments. Idea generation...
UC Berkeley Researchers Release CUA-Lite, an Open Platform Unifying Sandboxes, Data, Evaluation and RL for Computer-Use Agents
Source: MarkTechPost A team of researchers from UC Berkeley have released CUA-Lite, an open platform for computer-use agents...
Perplexity Details Its GPU Embedding Stack: How Ivy, Tulip and ROSE Serve pplx-embed
Source: MarkTechPost Retrieval quality in an AI search product is bounded by two things: how good the embedding...
GitHub Introduces Project HydraFusion: Runtime Multi-Model Orchestration That Builds a Workflow Per Coding Task in Copilot CLI
Source: MarkTechPost GitHub has released Project HydraFusion, a research preview that stops treating model choice as a one-time...
Adaption Labs Introduces ‘Invent a Dataset’: Training Data Generated From a Task Description, Not a Seed Corpus
Source: MarkTechPost This week, Adaption Labs released Invent a Dataset. The feature generates a structured, training-ready dataset from...