
Ai2 announces Bwen and Blama byte model checkpoints
Ai2 announces Qwen and Llama based byte model checkpoints alongside its Nature paper, with Stage 1 models for research.
Accessibility Adjustments
Use these optional tools to adjust reading and display preferences. These tools cannot resolve every accessibility barrier. Please contact the website owner if you need assistance.
New AI model releases can change the options available to developers, businesses and everyday users. ByteForward follows major launches and model updates across language, reasoning, coding and multimodal AI. This category focuses on what a release adds, how access works and which capabilities are available at launch. API availability, subscription requirements, regional restrictions and rollout stages help explain who can use a new model. Pricing, context limits and published evaluations provide further context for comparing the release with existing options. Our broader AI models coverage tracks how those systems develop after the initial announcement. Explore generative media for updates focused on AI image, video and audio creation. Browse the stories below to understand the practical details behind a model launch before deciding whether it fits your needs.

Ai2 announces Qwen and Llama based byte model checkpoints alongside its Nature paper, with Stage 1 models for research.

Perplexity’s smaller retrieval model can query an index built with its larger counterpart.

Uratori returns evidence judgments without generating an answer.

The update adds image inputs and more detailed video segmentation.

The 740 million parameter release combines text, image, video and audio embeddings for local retrieval.

Nano Banana 2.1 is available in the Gemini API. Google’s current documentation lists no shutdown date for the older Nano Banana 2 model.

Liquid AI has enabled image inputs in its paid d1 API, expanding structured decisions to visual inspection and screenshot tasks.
Mistral’s new flagship has 1.05 trillion parameters and a million token context window. Downloadable weights are planned for later in October.

The download combines a reasoning core with specialist tools and needs a GPU with 80 GB of memory.

Early access is limited while Reflection completes safety evaluations.

The compact CPU model handles seven languages and clips up to 30 seconds.

Kolibri brings open weights, reasoning and tool calling, with a million token context ceiling and practical deployment limits

Fastino GLiDE combines adaptive reasoning with structured decisions. Its September 30 release brings new routing options and practical limits for developers.

Tavus is testing Griffin Lite with selected research participants. Its video conversation results come with important limits on access and evaluation.

HeyGen Video generates complete scenes with sound from prompts and references. October launch rates depend on resolution and whether video references are used.

Perplexity Decider 27B combines public model weights with a hosted API that returns probabilities for classification and routing tasks.

Voyage Rerank 3 and Lite reorder search results with updated models, familiar token prices and practical limits for long documents.

Strands Decider 2B selects options and scores inputs for AI agent workflows, with open weights and important limits on reasoning and confidence.

Cloudflare releases two open decision models for Workers AI, with typed probability outputs, visual inputs and different speed and quality tradeoffs.

Cohere Embed 5 pairs Pro and Fast retrieval models with a shared index, multimodal inputs and separate text and image pricing.

Decagon pairs Voice 3 with Chord and a duplex architecture for customer calls. Here is what changed and what its early results establish.

Suno Speech creates spoken narration and background music in one track. The public beta is available on web and mobile, with uneven accents and timing still possible.

Microsoft introduces live transcription and two voice models, with character pricing, gated cloning and important Azure public preview limits.

Black Forest Labs opens a dedicated FLUX 3 image endpoint with layout controls, reference editing and output up to 4K. Pricing and preservation limits matter.