Release notes #
Information about release notes of INFINI Pizza is provided here.
Latest #
v0.1 is pending — nothing has been released yet. Everything below is the full feature set heading into the first release, in active development.
v0.1 (pending) #
Hello World, Hello Pizza!The first release: a distributed, real-time, AI-native search engine.
Distributed core #
- Cluster catalog metadata management, synchronized through Raft
(
openraft, with a monoio runtime binding) - RPC over Cap’n Proto, zero-copy both ways
- Multi-region organization for large clusters
- Promotion hardening — WAL catch-up barrier, primary-less sweep, idempotent role swaps; membership self-fence closes the zombie-primary window
- Online rolling reshard — progress and events API, fast-drain orchestration, hardlink staging
- WAN-safe Raft timing knobs; join tokens and API keys persisted
through the catalog Raft;
network.advertisesplit from binding - Formal verification — the consistency-critical protocols (promotion, allocation, reshard, doc-id reservation, layered visibility, fencing) are model-checked in TLA+; counterexamples have caught 18+ real bugs
Realtime engine and storage #
- True real-time indexing — no refresh, no flush; documents are searchable the moment they are acked
- WAL durability, local and Kafka-backed
- FireV11 segment format with a compaction pipeline, maintenance states, and task-lane metrics; per-collection manual compaction API with batched jobs
- Partial field updates (
_update) — toggle operator, optimistic concurrency over RPC, replicated fast lane, inplace checkpoints, andinplacetyped columns that turn hot counter / flag updates into O(1) column writes - Trash / recycle-bin lifecycle with a sweeper
- Memory profile dump endpoint; deletion-ledger rebuild at mount (documents can no longer resurrect after restart)
Query engine #
- BM25 scoring with execution explain
- Lucene query syntax and the Elasticsearch QueryDSL surface
- Distributed query execution with cross-shard aggregation reduce
- Point-in-time (PIT) pinned snapshot reads, ES
pitAPI - Supported query types:
match,match_all,match_phrase,multi_match,term,terms,bool,range,prefix,fuzzy,wildcard,regexp,suffix,exists,span,query_string,collapse,nested,joinandtraverse(relations), agraphprojection block that renders any hit set as nodes and edges,geo_bounding_box,geo_distance,vector/multi_vector(kNN),hybrid,semantic, and streaming search - Supported field types: keyword, text, booleans, all fixed-width
signed/unsigned integers (
i8…u64), floats and doubles, dates, arrays, objects, dense and sparse vectors
AI-native search #
- Server-side embedding inference — text→vector at index and query time through any OpenAI-compatible embedding service
- AI services registry — embedding model entries with provider presets, credentials resolved through the node keystore
semanticquery — semantic search with zero field knowledge: no vector field, no dimensions, no model config in the query body- TurboQuant 4-bit quantized vector search, with compaction and rolling-reshard support end to end
sparse_vectorfields over HTTP, indexing through search
Elasticsearch compatibility #
- ES
retriever/knnhybrid syntax translated into the native DSL - Execute or reject, never skip — a strict stage gate and top-level key whitelist: unsupported ES syntax fails loudly instead of silently degrading
- Keyed writes replace in place, render ES-style
updated/200;?key_as_id=truesurfaces the user key as_id; delete-by-key included
Text analysis #
- Built-in standard and whitespace analyzers, configurable stopwords, CJK-aware tokenization, lowercase/uppercase normalizers
- Analysis workbench in the console — catalog, pipeline builder, stage-by-stage debugging with per-stage diffs
- 40+ language analyzer family (
contrib/analysis-*): ICU segmentation, jieba/SmartCN/IK/pinyin/stconvert for Chinese, lindera (kuromoji/nori) for Japanese and Korean, the European stemmer set, and more
Security #
- Two-plane role model — built-in roles, API token lifecycle with one-time reveal, namespace-scoped key minting
- Authn/authz middleware across the data plane and write path
Observability and operations #
- Search audit ring with
_nodeendpoints; request-id correlation middleware - Per-node ES-shaped stats API and golden-metric trends
- Node task manager — batch flush, priority gate, unified lane view
- Node-local keystore for secrets —
${keystore:}config references, CLI management, masked view - Runtime node settings with dynamic log level; process RSS watchdog; recovery-time API and concurrency controls
Web console #
- Redesigned home with live cluster telemetry, topology and monitoring views, a FIRE storage explorer with per-field cost treemap, WAL entry drill-down, tasks tabs with compaction controls, a settings editor, and an AI services panel
- Web CLI — curl → console DSL conversion on paste, multi-select Run-N batch execution with stacked results
Tools and playground #
- Pizza CLI — quick interaction with the Pizza server
- In-browser playground — the same engine compiled to WebAssembly,
mounting real
.firesegments over a deterministic teaching dataset; every tutorial query runs live ( try it) - Docs site with dark mode, tutorials, a full reference sweep, and headless smoke tests asserting documented examples against the dataset