Enterprise model training
Frontier-class AI you can own
Price-performance comparison for a custom BI agent over 20 Oracle EBS scenarios. Only the model changes.
| Model | Pass rate | Cost / 1K conv. |
|---|---|---|
| Claude Opus 4.8API only | 61.7% | $2,380 |
| PAIR-trained 35B model2×H100 you own | 55.9% | $67 |
| GLM-5.216×H100 | 54.4% | $260 |
| GPT-OSS-120B1×H100 | 30.0% | $10 |
Frontier APIs are smart but expensive, and every call leaves your network. Small models can be self‑hosted but don’t perform well without tuning. For one customer’s BI agent, PAIR trained a 35B model that reached 91% of Opus 4.8’s score at 1/35 of the cost, running on two H100s.
Retrieval Optimizer
Automatically selects the best retrieval stack for your data
The Retrieval Optimizer in GoodMem Cloud benchmarks embedder and reranker combinations — open-weight or API — on your real queries, then tells you which ones are statistically near-best. Evidence, not vibes.
Your queries, not a leaderboard Every candidate is scored on your own workload and held-out data.
Honest uncertainty 95% ranges and near-best odds, so you can trade a sliver of quality for cost or latency.
Keeps going As your data and traffic change, GoodMem re-tunes the stack — no labeling required.
- Embedder: OpenAI Embedding 3 LargeReranker: Voyage Rerank 2.5 Mean NDCG 0.794 (95% range 0.760–0.825); 100% chance of being near-best, 100% chance of being best.
- Embedder: OpenAI Embedding 3 SmallReranker: Voyage Rerank 2.5 Mean NDCG 0.787 (95% range 0.753–0.819); 100% chance of being near-best, 1% chance of being best.
- Embedder: OpenAI Embedding 3 LargeReranker: Jina Reranker V3 Mean NDCG 0.739 (95% range 0.699–0.775); 31% chance of being near-best, 0% chance of being best.
- Embedder: OpenAI Embedding 3 SmallReranker: None Mean NDCG 0.690 (95% range 0.653–0.724); 0% chance of being near-best, 0% chance of being best.
Beyond text
Every page, ready for multimodal models
Real documents aren’t plain text. The answer often lives in a table grid, a chart, or a drawing — exactly what text extraction flattens away. GoodMem renders every page to an image at ingestion, even from Word, PowerPoint, and Excel, so a multimodal LLM can read the page the way a person would.

Query
“Does a G136 grapple fit a 321D excavator?”
GoodMem retrieves the page
Extracted text
321D • •Two dots — but under which grapples? The columns are gone.
Page image
The grid survives: the 321D row, and every column header above it.
A multimodal LLM reads the page
Answer: No.
The 321D takes the G120 or G126 on a B linkage. Both G136 columns are shaded “No Match” on that row.
- DOCX
- PPTX
- XLSX
Page images for PDF and all three Office formats, Excel included — rendered in-process, in pure Java, with no LibreOffice in the deployment.
Read how GoodMem captures pagesTwo agents ask for the salary bands. Only one gets them.
Permissions are checked before anything is searched. If a key can’t read a space, the request fails outright, and nothing leaks from the spaces it can read.
“What’s the salary band for an L5 engineer?”
support-agent
scoped to Help Center
403
Permission denied
The key can’t read Compensation, so nothing was searched.
hr-assistant
scoped to Compensation
comp-bands-2026.xlsx · p. 3
| Engineering L4 | $148k – $176k |
|---|---|
| Engineering L5 | $182k – $214k |
| Engineering L6 | $221k – $262k |
- Stored in
- Your PostgreSQL with pgvector. No proprietary store to migrate out of.
- Runs in
- GoodMem Cloud, your VPC, or fully air-gapped.
- Logged
- Per your policy: who asked, with which key, what they asked, and what came back.
- Supply chain
- Signed SLSA Build L3 provenance · distroless, non-root server image.
- Certified
- ISO/IEC 27001.
Cloud or local. Your choice.
GoodMem keeps your pipeline portable
Use frontier APIs and the open model stack with the same GoodMem retrieval pipeline. Connect vLLM, TEI, Ollama, and OpenRouter alongside your hosted providers.
- OpenAIHosted API
- AWSAWS BedrockManaged models
- OllamaLocal models

GoodMem
One retrieval pipeline
“The ability to swap providers, mix local and cloud infrastructure, and keep the retrieval pipeline consistent across all of them is genuinely impressive.”
Flexible deployment
Build agents on your own terms
- Self-host GoodMem
Self-hosted
Production-ready, free for commercial use.
Free
What’s included:
- Embed and ship anywhere, royalty‑free
- Unlimited memories per node
- gRPC API and SDKs in five languages
- No Auto-Optimizer
- Fully managedStart free
GoodMem Cloud
Continuously optimized for your data, ready in seconds.
From $15/mo
What’s included:
- Auto-Optimizer fine-tunes your retrieval models
- We run, monitor, scale, and update it
- Usage-based pricing, no surprise overages
- 14-day free trial, no credit card
- Contact sales
Enterprise
For OEMs, regulated industries, and service providers.
Custom
What’s included:
- On-prem or in your VPC
- Auto-Optimizer on your own infrastructure
- Custom LLM training by PAIR
- Source code access and a 99.99% SLA
# Install goodmem
curl -s "https://get.goodmem.ai" | bash