OpenRouter tool search keeps large function libraries out of the prompt
Functions marked defer_loading are searched for when a model needs them, rather than sending every definition on every turn. OpenRouter says tool search adds no extra charge.
Functions marked defer_loading are searched for when a model needs them, rather than sending every definition on every turn. OpenRouter says tool search adds no extra charge.
OpenRouter’s Server Tools Marketplace also includes Advisor, Subagent, Web fetch, Image generation, Apply patch, and Datetime. Most run server-side during a request, with no client-side tool loop.
OpenRouter has introduced the Server Tools Marketplace, with tools for grounding agents that include web search APIs, Shell, and tool search.
OpenRouter now hosts Qwen3.8 Max Prime from Alibaba_Qwen, describing it as a higher-throughput variant of Qwen3.8 Max with the same 2.4T-parameter capabilities.
OpenRouter says the issue that took Space Bunny Alpha offline is resolved and the model is serving again.
OpenRouter lists Recraft V4.1 Flash at $0.007 per image and 1.3 seconds per image. It says the output still looks like Recraft for photography, portraits, complex scenes, logos, and early concepts.
Microsoft AI says MAI-Image-2.6 is available to try on Foundry and OpenRouter.
OpenRouter has removed Space Bunny Alpha while its provider works through an issue, and says the model will return once it is stable.
OpenRouter says the stealth model Space Bunny Alpha is free. This time, the provider does not train on prompts or completions.
OpenRouter describes Space Bunny Alpha as a flash model with fast inference and adjustable reasoning. It accepts text, image, and video input.
He says OpenRouter routes Batch API requests to the cheapest provider, and that callers can use /chat/completions, /responses, or /messages on all models. Requests can be polled for a full response without a files API, and bring-your-own-key works.
OpenRouter says GPT-6 Sol and GPT-6 Luna use Astra's clearer, shorter, lower-jargon answers and OpenAI's improved prompt caching, with 90% off cached input reads.
OpenAI reports that GPT-6 Luna scores 66.6% on DeepSWE v1.1, comparable to Claude Opus 5 and Fable 5 at medium effort for 93-96% less per task, and matches GPT-5.6 Sol on factuality at higher effort for about 1/100th the cost.
OpenRouter is offering OpenAI’s GPT-6 Sol at $2 per million input tokens and $10 per million output tokens, and GPT-6 Luna at $0.10 and $0.50. OpenRouter says both are half the price of their GPT-5.6 predecessors, and that each tops its predecessor’s best AutomationBench score at a fraction of the cost per task.
OpenRouter lists those three differences for developers moving off Opus 5, and its migration guide covers what changed and how to update prompts.
OpenRouter is offering AnthropicAI’s Claude Opus 5.5 with a 1 million-token context window at $4 per million input tokens and $20 per million output tokens, a rate it says is 20% lower per token than Opus 5.
Each batch appears in the Batches tab of OpenRouter logs with its model, provider, status, and cost. Inputs and results are kept for 30 days or until the batch is deleted.
Images and files must be public URLs, not base64. Audio and video input are not accepted, and OpenRouter’s own web plugin is not available in batch.
OpenRouter’s Batch API is available for work that does not need an immediate answer. It cuts per-token pricing in half on most models and returns results within 24 hours.
LLM requests can be classified automatically with typesafeai's Jev in OpenRouter Classifiers, and the results can be analyzed in Explore.