Scaling Test-Time Compute in Search Mode

Scaling Test-Time Compute in Search Mode
Introducing medium, high, and ultrahigh effort tiers to the Query Agent's Search Mode.

Introducing medium, high, and ultrahigh effort tiers to the Query Agent's Search Mode.

Why AI won’t replace creatives, and how it can remove friction from messy workflows, lost files, and creative processes.

When a Weaviate query is slow, the first question is where the time went. Query profiling returns a per-stage, per-shard timing breakdown, making query performance issues visible.

This release brings the HFresh disk-based vector index and the built-in MCP Server to general availability, rebuilds cluster-wide async replication to run from a single scheduler (on by default), and adds two previews: the Boost API and Nested Object Filtering.

Server-side batching, retries, the blobHash data type, and multimodal ingestion — what to use when, with code.

Weaviate Cloud is now free to start across the entire product suite.

Engram, Weaviate's managed memory and context service for agentic applications, is now generally available.

Weaviate Cloud now supports more granular role-based access control with new Editor and Viewer roles for improved security and organizational management.

Use Weaviate's built-in MCP server to give Claude Code, Cursor, and VS Code hybrid search over your codebase and docs. No glue code.

Tokenization makes or breaks hybrid search. See how Weaviate's accent folding, custom stopwords, and /v1/tokenize endpoint power multilingual BM25.

A Researcher's Perspective on Retrieval Quality in RAG Systems

This release introduces the built-in MCP Server, Extensible Tokenizers, Diversity Search (MMR), and Query Profiling as previews, along with Incremental Backups, Gemini audio support for multi2vec-google, and the new BlobHash property type.