Developers can route supported open-weight language models to Baseten from Hugging Face model pages and client libraries, using either provider credentials or Hugging Face billing.
AI News Bank
A continuously updated, source-transparent newsroom for AI news, deals, companies, learning, product changes, rumors, and expectations.
Everything AI, separated by evidence type.
News
Confirmed reporting and analysis across models, policy, safety, infrastructure, and society.
25Deals
Funding, M&A, strategic investments, capacity commitments, and market structure.
15Companies
Platform moves, pricing, partnerships, competitive strategy, and business model shifts.
16Learn
Research artifacts, explainers, resources, benchmarks, and practical learning coverage.
104Changes
Model releases, access changes, APIs, deprecations, product updates, and availability.
7Rumors
Clearly labeled reports and unconfirmed signals with source-quality caveats.
23Expectations
Forecasts, deadlines, upcoming releases, and next checkpoints — never framed as fact.
Newest sourced posts
The replay-based evaluation tests whether language models know when to scaffold and when to push students to reason, while its authors caution that simulated sessions do not measure learning.
The unified endpoint now leads Google's model-and-agent developer stack, adding managed sandboxes, background jobs and optional server-side state while the older API remains supported.
The August 7 change opens more everyday health, educational and clinical queries while continuing to reroute dual-use research requests to Opus 5.
The cloud-hosted browser runs in V8 isolates and is designed around AI-agent automation, offering developers a lighter alternative for some browser-based tasks.
The open-weight speech model adds Arabic, Korean and Brazilian Portuguese voices, plus a production NIM for teams deploying voice agents on their own infrastructure.
The Apache-2.0 model targets always-on agents on consumer hardware, combining tool use, long-horizon reasoning and interleaved text-and-image input.
The 30-billion-parameter open model targets specialised agent tasks, while an open routing library directs requests across mixed-model systems.
The voice stack separates real-time audio from tool calls, model delegation and context maintenance to avoid audible stalls.
Repository teams can launch documentation, error-investigation and follow-up workflows from issue or pull-request comments.
Paid-plan users can choose how much supported models reason for each delegated task, trading potential quality gains against token and credit use.
Existing users have a short export window, while deployed apps should continue running and AI-powered projects need a replacement inference provider.