Answer any agent question directly on Parquet.
Unified retrieval infrastructure for AI workloads.
trusted by
- United States Government
-
- Bazinga Labs
built by
The creators of OpenSearch and engineering leaders across LinkedIn, Google, & Amazon.
Keep one copy of your data, in your own bucket.
architecture → · integration → · guide: hybrid search on Parquet files →
Infino embeds functions in SQL so agents can express complex questions in a single query.
-
An agent asks a question
How manyA count. The agent wants a number, not a pile of log lines. SQL does this with COUNT. errors last nightFilters. Keep the error-level rows and the last 24 hours. mention disk fullKeyword search. Match those exact words in the log body. Today this is a call to a search engine. or look like a capacity incidentSemantic search. Match meaning even when the log never says disk full. Today this is a call to a vector database.? Which serviceA group on the same hits. SQL GROUP BY. Today a second warehouse query after the searches return. had them, and which team owns itA join. Match each service to its owner in a second table. Today the agent looks this up in another system and stitches the answers together.?
-
It hides six jobs
- countHow many
- groupWhich service
- joinwhich team owns it
- filtererrors last night
- keyword searchdisk full
- semantic searchcapacity incident
-
One SQL statement runs all six
SELECT o.team, h.service, count(*) AS hits FROM hybrid_search( 'logs', 'body', 'disk full', 'embedding', :q, 1000 ) AS h JOIN owners o ON o.service = h.service WHERE h.level = 'error' AND h.ts > now() - interval '24 hours' GROUP BY o.team, h.service ORDER BY hits DESC;
Query your object storage directly to reduce costs.
Elastic Cloud Hosted · Platinum
OpenSearch Service
Infino Cloud Pricing. methodology →
Infino is built for the latency and scale that agents need.
- Warm
- Cold
- Warm
- Cold
Lower is faster. Infino is submitted and awaiting publication. Infino’s results are single node.
Lower is faster. All runs on c6a.4xlarge.
Lower is faster. All runs on c6a.4xlarge.
- Search (top-k)
- Count
Lower is faster. Infino is submitted and awaiting publication. Infino trails Lucene by 19% on top-k search and leads it by 26% on counts, where it is also faster than Tantivy.
Ask one question and watch keyword, semantic and hybrid search answer it side by side.
ask a question
Ask a question about arXiv abstracts, or pick a suggestion.
this is the retrieval step of an agent, made visible: one call each way, and hybrid decides when keyword and meaning disagree. getting it otherwise takes a search stack, a vector stack, and glue.
three searches run at once
Keyword, semantic and hybrid, each timed end to end over one Parquet table. Click a lane to see the SQL it sent.
via sql table functions idle times are full round trips to Infino Cloud, network included · the engine alone is faster
see what each one found
Hover a dot to read the result. Click it to pull in its neighbours.
Specs in the cloud or on-prem.
- search
- full-text · vector · hybrid · NL
- index
- BM25 (PFOR-delta, FST) · HNSW · OPANN + Sq16
- engine
- Rust + Small Language Models (SLMs)
- language
- SQL (Apache DataFusion) · REST · Query DSL · MCP · NL
- storage
- object storage, S3 · GCS · Azure Blob · on-prem
- format
- Apache Parquet
- deploy
- on-prem · self-hosted cloud · hosted cloud
- core license
- Apache-2.0
- security
- SOC 2 Type II
- encryption
- TLS in transit · AES-256 at rest