Item detail
github.com

Chroma: Apache-2.0 Open-Source Search Infrastructure for AI (Dense + Sparse + Full-Text + Multi-Vector Search, Native MCP Server, Chroma Cloud / Enterprise)

Chroma: Apache-2.0 Open-Source Search Infrastructure for AI (Dense + Sparse + Full-Text + Multi-Vector Search, Native MCP Server, Chroma Cloud / Enterprise) is an AI project that RepoRadar is tracking in its Radar section, currently rated Gold tier with a 'try now' verdict. Its strongest signal is practical usefulness, scored 10.0 out of 10.

Score8.7
Popularity0.0
Risklow
TierGold
Score breakdown
Usefulness10.0
Novelty7.0
Momentum10.0
Maturity6.8
Open-source/build8.4
Evidence7.2
Workflow potential9.4
Setup ease8.8

Popularity is tracked separately. Support, ads, sponsorships, and tips never affect these signals.

Why it matters

Most AI / ML engineers + RAG developers building production search-backed LLM applications have been either stitching together multiple vector DBs (Pinecone for dense + Elasticsearch for full-text + custom BM25 + custom reranker + custom metadata filtering -- high maintenance burden, no unified API), adopting a single-vendor vector DB (Pinecone, Weaviate, Qdrant Cloud) that locks-in the

Who should use it

Builders Power users

Who should skip it

Skip Chroma: Apache-2.0 Open-Source Search Infrastructure for AI (Dense + Sparse + Full-Text + Multi-Vector Search, Native MCP Server, Chroma Cloud / Enterprise) if the source link, documentation, or setup requirements do not align with your current workflow or stack.

About this signal

Chroma: Apache-2.0 Open-Source Search Infrastructure for AI (Dense + Sparse + Full-Text + Multi-Vector Search, Native MCP Server, Chroma Cloud / Enterprise) is tracked by RepoRadar as an AI project in the Radar section. First seen 2026-07-08; the source record was last checked on 2026-07-08. The current verdict is 'try now' with a Gold tier and easy setup difficulty. Chroma: Apache-2.0 Open-Source Search Infrastructure for AI (Dense + Sparse + Full-Text + Multi-Vector Search, Native MCP Server, Chroma Cloud / Enterprise) leads on practical usefulness (10.0) and momentum (10.0); its lowest signal is maturity (6.8), so factor that in before investing setup time. This page summarizes the public evidence on the linked source page and states where additional review is still needed.

How this item is evaluated

The Chroma: Apache-2.0 Open-Source Search Infrastructure for AI (Dense + Sparse + Full-Text + Multi-Vector Search, Native MCP Server, Chroma Cloud / Enterprise) record combines a 8.7/10 composite score with separate popularity (0.0), risk (low), and setup (easy) signals. See the scoring methodology for the current weights and evidence definitions.

Putting this into practice? Read How to vet an AI agent or MCP server before you wire it in for the checklist behind this score.

Risk explanation

The 28; 737* repo is at active maintenance but the consumer SHOULD note the dense + sparse + full-text + multi-vector search in one API is opinionated about storage backends -- the consumer SHOULD verify the embedded DuckDB / SQLite / ClickHouse storage backend matches their target scale before adopting; the consumer SHOULD note the late-interaction / SPLADE / ColBERT rerankers require embedding model selection + index configuration -- the consumer SHOULD verify their target embedding model + reranker configuration matches their use case before adopting; the consumer SHOULD note the SQL-style metadata filtering requires index configuration for high-cardinality fields -- the consumer SHOULD verify their target metadata schema matches the supported operators before adopting.

Evidence links
Closest alternatives / related signals
open-source apache-2-0 chroma search-infrastructure embedding-database vector-search dense-search sparse-search
Verification record

What RepoRadar actually verified

Tested in a bounded workflow

Bounded representative workflow retained by RepoRadar verification harness. Last checked 2026-07-13T10:35:18.897881Z.

passed · cohort-20260712-chroma-local-vector

Tester
RepoRadar automated local verification harness
Started
2026-07-13T10:35:12.405561Z
Completed
2026-07-13T10:35:18.897881Z
Environment
Windows 10 AMD64; Python 3.11.9; credential-stripped child environment; disposable home/cache
Install/setup time
1 minute(s)
Evidence scope
Bounded representative workflow
Cleanup
Per-check temporary home and work directory removed. Shared cohort package cache removed.
Actions exercised
  • Created a disposable home, work directory, and isolated package cache with credential-like environment variables excluded.
  • Created an ephemeral runbooks collection and inserted three documents with distinct vectors and team metadata.
  • Queried with a supplied incident vector, requested documents, metadata, and distances, and asserted the rollback runbook ranked first.
  • Executed bounded check: Insert three synthetic runbooks with supplied vectors and retrieve the expected nearest record plus metadata from Chroma.
  • Captured the complete sanitized stdout, stderr, exit status, artifact checks, and 6.49-second wall time.
Observed results
  • Command exited 0 after 6.49 seconds.
  • Chroma stored all three records and returned rollback as the nearest of two results with its platform-team metadata intact.
  • Expected marker 'CHECK_OK top_id=rollback top_team=platform count=3 returned=2' was observed in retained output.
  • Validated result.json: 3 required marker(s) present and 0 excluded marker(s) absent; size and SHA-256 are retained.
Observed strengths
  • The in-process client completed insert, count, similarity ranking, and metadata return in one credential-free workflow with an exact retained result marker.
Friction
  • The caller had to provide numeric embeddings explicitly; semantic embedding quality and model setup were intentionally outside this run.
  • Setup or runtime emitted 17 stderr line(s); the complete warnings/errors are preserved in the retained log.
Limitations
  • The run uses an ephemeral in-process collection and caller-supplied vectors; it does not cover an embedding function, persistence, server/cloud mode, filtering, concurrent writers, or production scale.
  • This credential-free disposable workflow does not establish production scale, model quality, reliability under sustained use, or team adoption.

Pricing assessment: The ephemeral local Chroma client required no Chroma Cloud plan and invoked no paid embedding model.

Privacy assessment: Only synthetic runbooks and supplied numeric vectors existed in process memory; no Chroma server, cloud endpoint, or embedding API was contacted.

Open retained test log →

Verification sources

Longitudinal intelligence

How this decision record is moving

Raw history JSON →

29 dated snapshots retained from 2026-07-08 through 2026-08-13; see the snapshot index for explicit coverage gaps. Stars, version, release, pricing, integration, risk, maintenance, verdict, score, and momentum fields remain explicit even when a source has not reported them. Repository momentum is a normalized 0–10 RepoRadar signal; GitHub stars appear only where the popularity monitor retained exact timestamped observations.

RepoRadar score8.7 current · +0.0 net
Repository momentum9.6 current · -0.4 net
GitHub stars (observed)29,052 current · +271 net
GitHub stars29,052 exact observation
Versionlatest
Last release2025-04-01T23:29:39Z
Maintenanceactive
Current risklow
Current verdicttry now
Pricing baselineNo structured commercial pricing baseline
Pricing checkedNot applicable or not recorded
Pricing freshnessNo dated commercial pricing review
Integrations baselineNo structured integrations recorded

Recent dated points

DateScoreMomentumStarsRiskVerdictMaintenance
2026-08-138.79.629,052lowtry nowactive
2026-08-128.79.629,038lowtry nowactive
2026-08-118.79.629,017lowtry nowactive
2026-08-108.79.629,004lowtry nowactive
2026-08-098.79.328,989lowtry nowactive
2026-08-088.79.328,979lowtry nowactive
2026-08-078.79.628,936lowtry nowactive
2026-08-068.710.0Not recordedlowtry nownot recorded
2026-08-058.710.0Not recordedlowtry nownot recorded
2026-08-048.79.628,936lowtry nowactive
2026-08-038.79.628,936lowtry nowactive
2026-08-028.79.328,923lowtry nowactive

Why the record changed

stars changed

Stars changed: 29038 → 29052.

stars changed

Stars changed: 29017 → 29038.

stars changed

Stars changed: 29004 → 29017.

stars changed

Stars changed: 28989 → 29004.

stars changed

Stars changed: 28979 → 28989.

stars changed

Stars changed: 28936 → 28979.

stars changed

Stars changed: 28923 → 28936.

stars changed

Stars changed: 28918 → 28923.

stars changed

Source-observed stars changed: 28915 → 28918. This reports the retained observation delta and does not infer why the upstream change occurred.

stars changed

Source-observed stars changed: 28914 → 28915. This reports the retained observation delta and does not infer why the upstream change occurred.

stars changed

Stars changed: 28896 → 28900.

stars changed

Source-observed stars changed: 28892 → 28896. This reports the retained observation delta and does not infer why the upstream change occurred.