FeedSourcesModels
PrivacyImpressum
⌘K
BB

Baseten Blog

49 articles total

Go to source

Stories, updates, and other resources from Baseten.

Go to source
  •  Announcing our partnership with OpenAI
  •  Securing the open frontier with NVIDIA OpenShell and Blaxel sandboxes
  •  Sheila Vashee joins Baseten as CMO
  •  Fine-tune on your LangSmith traces with Baseten Loops
  •  NVIDIA Nemotron 3 Diarization: real-time speaker labels at a cent per audio hour
  •  Introducing Baseten Hosted Tools
  •  LangChain trains custom models for LangSmith Engine with Baseten Loops
  •  DeepSeek-V4.1-Flash: more efficient prefill for coding agents
  •  Blaxel is joining Baseten to build the future of agentic infrastructure
  •  How Baseten makes pyannote’s diarization models 9.6x faster
  •  Optimizing delta weight syncs for managed rollouts
  •  Baseten leads Coval’s voice AI benchmark
  •  New MCP and skill for coding agents to use Baseten
  •  Best open-source models for post-training
  •  The efficient frontier of LLM inference
  •  Agentic kernels in production
  •  GLM 5.3: Scaling with post-training, intuitively explained
  •  The two AI gateway patterns in production inference
  •  How to run any open model inside DeepSeek Harness
  •  How leading platforms ensure observability for LLM inference