Google Research Releases ToolGrad: Answer-First Framework Hits 99.8% Pass Rate for Tool-Use Data Generation
Source: MarkTechPost Training an LLM to call tools reliably requires datasets that pair user queries with correct tool-use...
Meet Redis LangCache: A Managed Semantic Cache That Cuts LLM API Costs by Up to 90% and Returns Cache Hits Up to 15x Faster
Source: MarkTechPost Production LLM applications rarely receive a question nobody has asked before. Support assistants and RAG pipelines...
NVIDIA Details BioNeMo Inference Runtime (BioIR): 2.90x Higher Boltz-2 Folding Throughput and 58.5K Residues per GPU-Hour on 8xH100
Source: MarkTechPost Biomolecular structure prediction has shifted from single-target runs to proteome-scale worklists. The bottleneck is no longer...
OpenAI Launches the Agents API in Public Beta, Putting the Codex Harness Behind One API Call
Source: MarkTechPost OpenAI has released the Agents API in public beta. It gives developers the same harness and...
DeepSeek AI Released DeepSeek-V4.1-Flash with 1M Context, FP4 KV Cache, and Cross-Layer Attention Reuse
Source: MarkTechPost Long-horizon agents have turned LLM serving into an input-heavy workload. Repeated prefills and million-token contexts leave...
LandingAI Releases Agentic Document Extraction Gen2 with DPT-3 Pro and DPT-3 Verity
Source: MarkTechPost LandingAI has shipped Agentic Document Extraction (ADE) Gen2, a rebuild of its document intelligence stack around...
Google Open-Sources Mantis: A Modular Skills Toolkit That Lets Coding Agents Find, Reproduce and Patch Vulnerabilities
Source: MarkTechPost Google has open-sourced Mantis, a stack-agnostic toolkit of security review skills that lets an AI coding...
Meta Introduces Muse, a Personal AI Agent That Runs on Its Own Dedicated Secure Cloud Computer
Source: MarkTechPost Today, Meta has introduced Muse, a personal AI agent that takes actions rather than just answering...
NVIDIA Announces CUDA Rust with cuda-oxide (SIMT) and cutile-rs (Tile) for Compile-Time-Safe GPU Kernels
Source: MarkTechPost NVIDIA has announced CUDA Rust, a push to make Rust a first-class language for writing GPU...
Google DeepMind Releases AlphaGenome Atlas With Precomputed Molecular Effect Predictions and AVI Scores for 9 Billion Human DNA Variants
Source: MarkTechPost Google DeepMind has released AlphaGenome Atlas, a catalogue of precomputed predictions for the molecular effects of...