Latest AI and tech news

runpod.io
RAG · GPUs & chips +9 C2BA7

Selecting the right serving engine for your embedding model can dramatically outperform hardware upgrades, yielding up to an 11x throughput increase on the same GPU.
runpod.io
GPUs & chips · Data centers +8 0BD73

When a private GPU pool beats on-demand GPUs, how reserved capacity is billed, what happens when traffic bursts above the pool, and what usage data to bring before you commit.
runpod.io
Inference · AI infrastructure +9 48F15

Qwen3.8-Flash-Next needs vLLM 0.29, so the Hub's one-click path won't serve it yet. Here are the validated flags, hardware math, and cold-start numbers for running it on Runpod.
runpod.io
AI · MCP +9 12A5B

Learn how to optimize Model Context Protocol (MCP) tools to prevent oversized responses from exhausting your AI agent's context window.
runpod.io
AI infrastructure · Languages +8 1B06E

Learn how to edit videos with Pruna P-Video-Edit using Runpod’s public endpoint, Playground, curl, and Python.
runpod.io
Cybersecurity · AI infrastructure +8 D85CB

The certification audit closed with zero findings, giving international customers a standing answer instead of a bespoke security questionnaire.
runpod.io
Cost of AI · cheaper +7 45352

Renting a model means your millionth request costs what your first one did. Owning the weights makes cost per request something your engineers can lower.
runpod.io
AI infrastructure · Kubernetes +8 71039

How to get up and running with Kimi K3 without the overhead of running an entire cluster.
runpod.io
AI agents · AI infrastructure +11 0BE72

With more than half of CLI usage driven by agents, there's strong demand for Runpod's agent-ready infrastructure.
runpod.io
GPUs & chips · Cloud +10 C3457

Runpod Serverless now scores fallback GPU types so workers can land on compatible capacity when your first choice is contended.
runpod.io
GPUs & chips · AI infrastructure +8 76AD5

Optimize multi-node GPU cluster performance and cost-efficiency by aligning parallelism strategies, network fabric selection, and performance benchmarking.
runpod.io
AI coding tools · AI infrastructure +10 FC608

Qwen 3.8 is punching well above its weight for its relatively small size compared to big foundational models. Is there a such thing as a free lunch?
runpod.io
Cost of AI · GPUs & chips +9 1B501

Our pricing philosophy in one line: we move prices to keep GPUs available.
runpod.io
Languages · Cloud +12 F4572

Deploy Qwen3.8-27B on Runpod Serverless with vLLM, then send your first request with curl and the OpenAI Python SDK.
runpod.io
GPUs & chips · Inference +10 CACFA

vLLM cold start optimization on Runpod Serverless: compile cache, weight prefetch, CUDA graph config
runpod.io
Languages · Cybersecurity +9 1CD81

Python versions carry a security clock and a performance upgrade you're leaving on the table — here's what changes when you move off an old interpreter, and when it's fine to wait.
runpod.io
AI infrastructure · optimizing +7 5E420

By optimizing test architecture through pool-mode parallelization and isolated worker identities rather than increasing compute resources, the team successfully reduced merge-queue CI times by 55%.
runpod.io
GPUs & chips · explained +2 A9FA9

runpod.io
AI · six +3 6A6AD

runpod.io
AI infrastructure · storage +3 0D529

runpod.io
AI · AI infrastructure +3 4BE4E

runpod.io
Robotics · Local models +7 3B5C9

runpod.io
AI research · GPUs & chips +9 421CC

runpod.io
AI infrastructure · them +4 6140B

runpod.io
AI infrastructure · fast +5 0513D

runpod.io
Local models · open-weight +4 DC6B1

runpod.io
AI infrastructure · endpoint +3 FCA02

runpod.io
Data centers · Cloud +5 BE088

runpod.io
LLMs · control +3 1AC44

runpod.io
GPUs & chips · AI +5 9E869

runpod.io
MCP · GPUs & chips +6 6B0A1

runpod.io
AI assistants · AI infrastructure +10 A0C65

runpod.io
AI infrastructure · source +6 268D7

runpod.io
Inference · cold +2 FEAD6

runpod.io
Automation · GPUs & chips +11 64D1E

runpod.io
AI infrastructure · flash +3 46BB9

runpod.io
Inference · AI infrastructure +5 61AE0

runpod.io
Inference · AI infrastructure +4 EE318

runpod.io
Cloud · Data engineering +11 E757E

runpod.io
Enterprise AI · GPUs & chips +8 D7C1B

runpod.io
Cloud · AI infrastructure +4 AEBE1

runpod.io
available · deploy D0F63

Keyboard shortcuts

On. Switch them off if they collide with your assistive tools; ? still opens this sheet.

Go to

Press g then the letter.

  • gh Latest
  • gs Sources
  • gm Media
  • gv Videos
  • gp Podcasts
  • gc Calendar
  • gd Decoder
  • gz Dataviz
  • ga Datasets
  • gb Blog
  • gk Markets
  • gj Careers
  • gn Prompt Notebook

On this page

  • / Focus search, where there is one
  • t Back to top
  • ? This list
  • Esc Close