LLMs & Models Articles
Browse 579 articles about LLMs & Models.

What Is Local AI Inference? Why NVIDIA RTX Spark Changes Everything
NVIDIA's RTX Spark chip brings 128GB unified compute to laptops, enabling large LLMs to run locally without internet. Here's what it means for AI builders.

MAI Transcribe 1.5: Is Microsoft's New Model Really the Best Transcription AI?
MAI Transcribe 1.5 claims to be the world's most accurate and fastest transcription model—5x faster than competitors. Here's what the benchmarks show.

Minimax M3: A 1M Token Context Coding Model That Claims to Beat GPT 5.5
Minimax M3 is a coding model with a 1 million token context window that outperforms GPT 5.5 on SWE-bench Pro. Here's what it can do and how to access it.

NVIDIA Nemotron 3 Ultra: 550B Parameters, 5x Faster, 30% Cheaper for Agents
NVIDIA's Nemotron 3 Ultra is a 550B open-weight model built for agentic tasks. It beats trillion-parameter models on agent benchmarks at a fraction of the cost.

What Is Multi-Tier On-Policy Distillation? How NVIDIA Trained Nemotron 3 Ultra
NVIDIA used multi-tier on-policy distillation to train Nemotron 3 Ultra. Learn how this technique produces stronger models than single-task training.

What Is NVIDIA Nemotron 3 Ultra? The 550B Open-Weight Model Built for Agents
NVIDIA Nemotron 3 Ultra is a 550B parameter open-weight model optimized for agentic tasks. Learn how it compares to frontier models and how to access it.

NVIDIA Nemotron 3 Ultra vs Claude Opus 4.8: Which Open Model Wins for Agents?
Compare NVIDIA Nemotron 3 Ultra and Claude Opus 4.8 on agent benchmarks, speed, cost, and tool-calling to find the right model for your agentic workflows.

What Is Claude Opus 4.8? Anthropic's Incremental Model Update Explained
Claude Opus 4.8 brings improved agentic task performance and a new /workflows command. Here's what changed, what didn't, and when to use it.

What Is Claude Opus 4.8 Overthinking? Why Max Mode Can Hurt Performance
Claude Opus 4.8 sometimes overthinks on constitutional questions in max mode, reducing effectiveness. Here's what it means and when to use high vs max.

What Is the Vending Bench? The AI Business Benchmark That Exposes Real-World Agent Gaps
Vending Bench tests how AI models run an actual business. Claude Opus 4.7 outperformed 4.8 on it—here's what that tells you about model selection.

Claude Opus 4.8 vs GPT 5.5 on Coding Benchmarks: What the DeepSuite Results Show
Compare Claude Opus 4.8 and GPT 5.5 on the DeepSuite software engineering benchmark. See which model wins on real coding tasks.

What Is the History of AI? From Alan Turing to Claude Code in 100 Years
Trace AI history from Turing's Bombe to the transformer revolution and Claude Code. Understand the breakthroughs that made modern AI agents possible.

What Is Arc AGI 3? How Claude Opus 4.8 Achieved State-of-the-Art Fluid Intelligence
Arc AGI 3 tests fluid intelligence in AI models. Claude Opus 4.8 reached 1.5% — the highest score ever — by reasoning at a higher abstraction level.

What Is Backpropagation? The Algorithm That Made Modern AI Agents Possible
Backpropagation solved the multi-layer neural network training problem in 1986. Learn how this algorithm underpins every LLM and AI agent today.

What Is NVIDIA Neotron 3 Ultra? The Open-Source AI Model That's 5x Faster
NVIDIA Neotron 3 Ultra is a 550B open-source model that's 5x faster and 30% cheaper than competing frontier models. Here's what it means.

What Is Claude Opus 4.8 Honesty Mode? How Anthropic's Model Flags Uncertainty
Claude Opus 4.8 improves honesty by flagging uncertainties and avoiding unsupported claims. Here's what changed and why it matters for AI agents.

How to Use NVIDIA Cosmos 3 to Generate Synthetic Training Data for Robotics
Cosmos 3 generates synthetic video data for training robot arms and physical AI systems. Learn how to run inference and what use cases it unlocks.

What Is NVIDIA Cosmos 3? The World Foundation Model for Robotics and Physical AI
NVIDIA Cosmos 3 is a multimodal world model that handles text, images, video, audio, and actions in one architecture. Here's what it means for AI builders.

How to Run Local AI on AMD: ROCm, LM Studio, Ollama, and ComfyUI Setup
AMD's ROCm platform now supports PyTorch, Ollama, LM Studio, and ComfyUI out of the box. Here's how to set up a full local AI stack on AMD hardware.

Claude Opus 4.8 vs Claude Opus 4.7: What Actually Changed?
Claude Opus 4.8 fixes 4.7's biggest complaints: less attitude, better honesty, and restored creativity. Here's a real-world comparison of both models.