Professor Emeritus Dimitri Bertsekas, influential computer scientist and prolific author, dies at 83
Source: MIT News – Artificial intelligence Dimitri Bertsekas PhD ’71, the Jerry McAfee (1940) Emeritus Professor in Engineering...
Unsloth vs Axolotl vs TRL vs LLaMA-Factory: A Fine-Tuning Framework Comparison on Speed, VRAM, and Multi-GPU
Source: MarkTechPost Four open source projects dominate LLM fine-tuning today. Unsloth, Axolotl, TRL, and LLaMA-Factory all wrap the...
Cisco Foundation AI Releases Antares: 350M and 1B Open-Weight Models That Localize Known Vulnerabilities Inside Real Codebases
Source: MarkTechPost Cisco Foundation AI has released Antares, a family of security small language models (SLMs) built for...
Poolside Releases Laguna S 2.1, an Open-Weight Agentic Coding Model Punching Above Its Weight Class on SWE-Bench Multilingual
Source: MarkTechPost Poolside has released Laguna S 2.1, a 118B-parameter open-weight model built for agentic coding. It is...
Google Releases Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber: A Cheaper, More Token-Efficient Flash Tier Built for Agentic Workloads
Source: MarkTechPost Developers building production agents need higher token efficiency, lower latency, and more reliable performance. Today, Google...
Validating Distributed LLM Serving Benchmarks with NVIDIA srt-slurm, SLURM Recipes, Parameter Sweeps, and Pareto Analysis
Source: MarkTechPost In this tutorial, we explore NVIDIA’s srt-slurm framework and learn how we use srtctl to convert...
Meta Open-Sources Astryx: An Agent-Ready React Design System With 150+ Accessible Components, Seven Themes, and a CLI
Source: MarkTechPost Meta has released Astryx, an open source design system that is fully customizable and built to...
NVIDIA Releases Cosmos 3 Edge: A 4B-Parameter Open World Model That Reasons and Generates Robot Actions On-Device
Source: MarkTechPost NVIDIA has released Cosmos 3 Edge, a 4-billion-parameter open world model built to run on-device. It...
Alibaba’s Tongyi Lab Releases Qwen-Audio-3.0-TTS, a Hosted Text-to-Speech Model in Flash and Plus Tiers Across 16 Languages
Source: MarkTechPost Alibaba’s Tongyi Lab has released Qwen-Audio-3.0-TTS, a production-oriented text-to-speech (TTS) system. The model ships in two...
Someone Fine-Tuned OpenBMB’s MiniCPM5-1B on Claude Fable 5 Traces to Ship a 657MB Local Thinking Model
Source: MarkTechPost A community developer, GnLOLot, has published a 1B model that runs fully on local hardware. The...