Ling 3.1 Flash launches free on Vercel AI Gateway
Hybrid reasoning model with 560B parameters and 262K-token context window available free through October 13, integrated into Claude Code, Codex, and fx via single API Gateway endpoint.
Consolidates model access across three coding agents through unified gateway routing, eliminating per-agent credential management. Developers can swap models at runtime without code changes using `/model` command.
Replaces separate agent API keys with single AI Gateway credential. Requires: `npx vercel ai-gateway setup` + model selection. Worth trying now if already using Vercel agents; standard model ID converts to paid after October 13, `-free` variant stops serving.
- “Ling 3.1 Flash is a hybrid reasoning language model with 560B total parameters and 25B active per token”
- “It has a 262K-token context window on AI Gateway”
- “The model is designed for coding, multi-step analysis, and agents that use tools”
- “The standard model ID is free during the promotion and begins billing when it ends”
vercel-ai-gatewayling-3.1-flashcoding-agentscontext-windowmodel-routing
AI Gateway adds Browserbase Search and Fetch tools
Browserbase Search and Fetch are now available through AI Gateway, letting you add web search and page retrieval to any model supporting tool calling with a single API key.
Eliminates vendor lock-in for web-aware agents—switch between model providers without rewriting tool logic. Reduces boilerplate for RAG patterns that need live web data.
Replaces building custom HTTP adapters for web search/fetch. Requires AI SDK 7.0.116+, Browserbase integration, and `AI_GATEWAY_API_KEY` setup. Ready to use now with tool calling models; check Browserbase pricing for quota/cost.
- “AI Gateway lets you use Browserbase's tools across model providers with one API key”
- “These helpers are available in AI SDK 7.0.116 and later”
- “Add web search and page retrieval to any model that supports tool calling, and keep the same tools when you switch models”
ai-gatewaytool-callingweb-searchragvercel
Google releases Agent Development Kit for Kotlin 1.0
Compile-time tool schema generation via KSP eliminates runtime reflection overhead; requireConfirmation guards high-impact operations on mobile.
Kotlin teams can now build agentic systems without Python detour, with native Android persistence (Room, AppSearch) and on-device inference (LiteRT-LM, ML Kit) integrated. Session serialization and context compaction reduce token waste in multi-turn workflows.
Replaces Python-based agent logic for JVM/Android stacks. Requires KSP compiler plugin and Kotlin 1.0+. Production-ready for on-device + hybrid setups; start with single resumable agent and explicit tool confirmation before multi-agent hierarchies.
- “completely agnostic to specific model backends, session providers, or memory systems”
- “handling tool schemas at compile time with KSP keeps startup fast on mobile targets”
- “Production readiness depends more on lifecycle recovery and deterministic tool boundaries than on agent count”
- “The ADK for Kotlin is open source and available on GitHub”
kotlinagentsandroidtool-callingon-device-ai
Worker Previews isolates branches in production-like environments
Each Git branch deploys to its own ephemeral Worker instance with isolated Durable Objects, persistent state, and observability—replacing shared staging environments and enabling agents to test autonomously before merge.
Eliminates staging/production parity gaps by letting developers and agents verify behavior in production-grade conditions per branch without contention. Durable Object isolation per Preview prevents state collisions when running concurrent tests, critical for agents making atomic changes at scale.
Replaces manual staging environment management and shared test databases. Requires Cloudflare Workers, Wrangler CLI, and a previews block in wrangler.json with base configuration. Ready now—native Cloudflare feature; direct lift for existing Worker projects. Agents can integrate via MCP servers (Browser Run, Workers Observability) for headless testing and trace inspection.
- “Each Git branch gets its own isolated environment and URL”
- “every time you run npx wrangler preview, Cloudflare automatically creates a new Durable Object namespace and Container application for that Preview”
- “Each Preview runs as a real version of your Worker”
- “Every Preview has its own isolated and persistent state, with Durable Objects and Containers”
cloudflare-workerspreview-environmentsstaging-alternativeagent-workflowsdurable-objects
AI Gateway adds TypeSafe client and HTTP API support
Route Jev probabilistic decisions through AI Gateway via TypeSafe client, HTTP API, or AI SDK—unified billing and observability across all three paths.
Eliminates integration friction for teams already using TypeSafe clients; enables language-agnostic access to typed decision models without rebuilding evaluation logic. Single gateway for observability means model calls, costs, and latency appear alongside your LLM usage.
Replaces separate TypeSafe and AI Gateway integrations with unified routing. Requires only base URL and API key swap for existing TypeSafe clients; zero code changes to evaluation calls. HTTP API path adds zero friction for non-TypeScript stacks. Worth trying now if you're on TypeSafe or evaluating probabilistic decision models for agent control flow.
- “Point an existing TypeSafe client at AI Gateway without changing its evaluation calls.”
- “State goes in, and typed answers come out with probabilities attached, so there's no generated text to parse.”
- “Requests are billed through AI Gateway on all three paths, so they appear alongside your other model calls in usage and observability.”
ai-gatewaytypesafeevaluationhttp-apiobservability