News
AI Provider Index news.
Updates, product notes and practical coverage for AI platforms and operators.
Open Baseten Adds Hosted Inkling Small Model API Access
01 Aug 2026 · Ronald
Baseten Adds Hosted Inkling Small Model API Access
Baseten has added Inkling Small to its hosted Model APIs and dedicated-deployment service. The open-weight mixture-of-experts model has 276 billion total parameters, 12 billion active parameters, a one-million-token context window and native text, image and audio inputs.
Open Featherless adds hosted Kimi K3 access with vision support
30 Jul 2026 · Ronald
Featherless adds hosted Kimi K3 access with vision support
Featherless has added hosted access to Moonshot AI's open-weight Kimi K3 model through its OpenAI-compatible API. The initial configuration offers vision input, FP8 quantisation and a 32K context window, while the provider works towards larger context options.
Open Liquid AI launches compact LFM2.5 encoders for edge use
29 Jul 2026 · Ronald
Liquid AI launches compact LFM2.5 encoders for edge use
Liquid AI has released two open-weight LFM2.5 encoder models for classification, retrieval and token-level language tasks. The compact 230-million and 350-million parameter models target local, on-premises and high-volume deployments, with an 8,192-token context window and public evaluation code.
Open Anthropic Launches Claude Opus 5 for Long-Running Agents
25 Jul 2026 · Ronald
Anthropic Launches Claude Opus 5 for Long-Running Agents
Anthropic has released Claude Opus 5 across Claude and its API, positioning the model for demanding coding, research and agent workflows. The launch keeps Opus 4.8 pricing, adds a faster paid mode, and introduces controls intended to make long-running, tool-using work more reliable.
Open xAI Launches Grok 4.5 for Coding and Agentic Work
16 Jul 2026 · Ronald
xAI Launches Grok 4.5 for Coding and Agentic Work
xAI has released Grok 4.5 for coding, agentic tasks and knowledge work, with immediate access through its API, Grok Build and Cursor. The company is positioning the model around engineering performance, faster generation and lower token use, while its benchmark and efficiency claims still require independent testing.