LLMs
Large language and foundation models, the labs that train them and the releases that move the field
Latest stories See all →
- Why We Post-Trained Our Own Reasoning Model
- AI Agent Benchmark Results 2026: Why Your Computer Use Tool Is Garbage
- LLM UI Testing: Smarter Interface Testing with AI [2026]
- Mistral and Mozilla are bringing open, private and multilingual AI to your web browser
- Salesforce Expands Missionforce with Purpose-Built AI and New Partnership with OpenAI to Advance Mission Intelligence
- Build a live RAG pipeline with Apify, n8n, and Qdrant
- The Best Engineer I Ever Worked With Wrote the Least Code
- How Computer Use Actually Works: Inside OpenAI, Anthropic, and Google’s AI Agents
- RAG vs Fine-Tuning: Which One Do You Actually Need?
- I’m not an engineer. I fine-tuned my own language model on a MacBook Air
- GPT-6 Astra Is So Good, It Changes the Rules of Software Development
- How to Connect Blink to Claude: Build and Ship Apps from a Chat
- The Hugging Face Wake-Up Call: What an Autonomous AI Attack Means for CISOs
- Claude Data Privacy: How Anthropic Handles Your Data (2026)
- Does Claude Train on Your Data? (2026 Facts)
- Best Prompt Analytics and Performance Tools for Engineering Teams (2026)
- OpenArt Arena's AI Model Leaderboard
- OpenAI president: “The computer should be there to empower you.” So stop retooling software for AI agents
- Graphs, knowledge graphs, & context graphs: Which one do you need?
- Real-Time Evidence Extraction for Risk Adjustment: What the NLP Infrastructure Requires
- Meta lets Claude and Codex configure WhatsApp Business via MCP
- OpenAI’s voice model doesn’t think. That’s the point.
- Salesforce and OpenAI Customers Are Putting AI to Work
- Knowledge Graph Integration With Poro2 For Enriching Medical Text Processing
Videos
- Gemini 3.8 Live vs Extended Thinking: Which one should you choose?
- DeepSeek and Kimi Secretly Sent Your Prompts to Claude
- DeepSeek V4.1 Flash: The New Speed King That Also Thinks Straight
- Can AI Agents Automate LLM Post-Training? PostTrainBench Results | Maksym Andriushchenko
- [Live Session] Choosing safer LLMs: From LLM benchmarks to your production agents
- This Open-Weights TTS Beat ElevenLabs... Then I Read the License
- ClickStack Demo Day (2026-08-28) - LLM observability out of the box
- OmegaClaw: An AI Agent Built on Symbolic Logic, Not Just an LLM
- Gemini Omni Flash 1.1 Is Here — Here’s What’s NEW
- Gemini 3.8 Flash: The model no one expected!
- Agentic approaches to processing long videos with Gemini
- Agentic video understanding in Gemini
Podcasts
Recent episodes
- Firmware Analysis, Linux Malware, Future of AI - BTS #82
- AI safety requires action, not promises
- Risky Business #853 -- We're all gonna die, apparently
- SN 1096: Are we the Krell? - 153 Million Driver's Licenses Leaked
- Ep. 031 - EMERGENCY EPISODE: Are We Doomed? | Jordan Nanos, Doug O'Laughlin, Max Kan, Joey Brookhart
- What Context Really Means in Data Engineering and AI
- Trump Says: No Slowdown!
- EP 192: Talking Apple Silicon with Apple's Tom Boger, Kaiann Drance, and Sri Santhanam
- Superhuman acquired notetaker Fathom; plus, OpenAI bought smartphone camera maker Glass Imaging, and AI infrastructure company Cornelis raised $205M
- 1027: Building an Always-On AI Agent for Busy Parents, with Dr. Dilani Kahawala
- 1027: Building an Always-On AI Agent for Busy Parents, with Dr. Dilani Kahawala
- Why AI models are obsessed with creatures