Answer any agent question on Parquet.
Scalable search + analytics infrastructure on object storage, for 10x cheaper.
trusted by
- United States Government
-
- Bazinga Labs
built by
The creators of OpenSearch and engineering leaders across LinkedIn, Google, & Amazon.
Keep one copy of your data, in your own bucket.
architecture → · integration → · guide: hybrid search on Parquet files →
Infino embeds functions in SQL so agents can express complex questions in a single query.
-
An agent asks a question
How manyA count. The agent wants a number, not a pile of log lines. SQL does this with COUNT. errors last nightFilters. Keep the error-level rows and the last 24 hours. mention disk fullKeyword search. Match those exact words in the log body. Today this is a call to a search engine. or look like a capacity incidentSemantic search. Match meaning even when the log never says disk full. Today this is a call to a vector database., and are tied to last week’s outageGraph search. Keep the errors linked to the outage’s record through a key they share, like a host or a service, even when they never mention it. Today this is a call to a graph database.? Which serviceA group on the same hits. SQL GROUP BY. Today a second warehouse query after the searches return. had them, and which team owns itA join. Match each service to its owner in a second table. Today the agent looks this up in another system and stitches the answers together.?
-
It hides seven jobs
- countHow many
- groupWhich service
- joinwhich team owns it
- filtererrors last night
- keyword searchdisk full
- semantic searchcapacity incident
- graph searchtied to last week’s outage
-
One SQL statement runs all seven
SELECT o.team, h.service, count(*) AS hits FROM hybrid_search( 'logs', 'body', 'disk full', 'embedding', :q, 1000 ) AS h JOIN graph_walk('logs', 'incidents', [4521], 2, 5000) AS g ON g._id = h._id JOIN owners o ON o.service = h.service WHERE h.level = 'error' AND h.ts > now() - interval '24 hours' GROUP BY o.team, h.service ORDER BY hits DESC;
Combine modes to increase retrieval accuracy.
- keyword searcha row’s words
- semantic searchits meaning
- graph searchits links
- 100%retrieval accuracy
Retrieval is cheap and scalable.
Elastic Cloud Hosted · Platinum
OpenSearch Service
Infino Cloud Pricing. methodology →
Infino is built for latency from the ground up.
- Warm
- Cold
Internal 10M-document run. Its full results are not published yet. The published engine benchmarks, with a way to reproduce each one, are on the benchmarks page.
- Warm
- Cold
Internal 10M-document run. Its full results are not published yet. The published engine benchmarks, with a way to reproduce each one, are on the benchmarks page.
Internal 10M-document run. Its full results are not published yet. The published engine benchmarks, with a way to reproduce each one, are on the benchmarks page.
Lower is faster. Infino is submitted and awaiting publication. Infino’s results are single node.
Lower is faster. All runs on c6a.4xlarge.
Lower is faster. All runs on c6a.4xlarge.
- Search (top-k)
- Count
Lower is faster. Infino is submitted and awaiting publication. Infino trails Lucene by 19% on top-k search and leads it by 26% on counts, where it is also faster than Tantivy.
Ask one question and watch keyword, semantic and hybrid search answer it side by side.
ask a question
Ask a question about arXiv abstracts, or pick a suggestion.
this is the retrieval step of an agent, made visible: one call each way, and hybrid decides when keyword and meaning disagree. getting it otherwise takes a search stack, a vector stack, and glue.
three searches run at once
Keyword, semantic and hybrid, each timed end to end over one Parquet table. Click a lane to see the SQL it sent.
via sql table functions idle times are full round trips to Infino Cloud, network included · the engine alone is faster
see what each one found
Hover a dot to read the result. Click it to pull in its neighbours.
Specs in the cloud or on-prem.
- search
- full-text · vector · hybrid · graph · NL
- index
- BM25 (PFOR-delta, FST) · HNSW · OPANN + Sq16
- engine
- Rust + Small Language Models (SLMs)
- language
- SQL (Apache DataFusion) · REST · Query DSL · MCP · NL
- storage
- object storage, S3 · GCS · Azure Blob · on-prem
- format
- Apache Parquet
- deploy
- on-prem · self-hosted cloud · hosted cloud
- core license
- Apache-2.0
- security
- SOC 2 Type II
- encryption
- TLS in transit · AES-256 at rest