Websites, AI agents, LLM apps, RAG & search, and voice agents — built end to end for startups and small agencies. Open any build to read the case study.
44 builds in our portfolio.
Copperleaf Goods · Support agents
Design and build a modern e-commerce website with an embedded AI shopping assistant: natural-language product discovery, grounded answers from the catalog and policies, and checkout nudges that lift conversion.
What we built
A full e-commerce site with a built-in shopping agent that recommends products, answers sizing and returns questions, and recovers carts — the whole storefront, not a bolt-on widget.
Cedar Point Health · Voice agents
Build a HIPAA-conscious patient portal with an intake agent: appointment booking, structured history collection over chat or voice, and clinician-ready summaries — with strict guardrails and human escalation.
What we built
A patient-facing website where people book, fill intake, and ask questions — with a guarded agent that collects history and hands a clean summary to the care team.
Hartwell & Cross LLP · RAG & Search
Build a law-firm website with two agents: a client-intake agent that qualifies matters and books consultations, and a case-search agent that answers questions over the firm's document corpus with citations.
What we built
A polished firm website with a front-of-site intake agent that qualifies matters and books consults, plus an internal case-search agent that answers over documents with citations — marketing site and legal RAG in one build.
SkyHunter Partners · Workflow automation
Join the SkyHunter partner pool and pick up overflow agent builds from our client pipeline. A few projects a month, fully remote across the US, LATAM, and Europe.
What we built
We route vetted agent-build projects to shops with capacity — real paid client work, not spec. Take what fits your stack.
Verity Labs · Support agents
Build a support agent that lives in their onboarding flow, resolves setup questions against product docs, and escalates cleanly. Ship it inside their web widget.
What we built
An in-app agent that walks new users through setup, answers product questions from the docs, and books a human call when it's stuck.
Northlight AI · RAG & Search
Build a retrieval agent over a large, versioned HR policy corpus with strict grounding and jurisdiction awareness. Accuracy and citations are the whole point.
What we built
Employees ask 'how much parental leave do I get in Ohio?' and get the exact, cited policy — not a wrong guess.
Aegis Research · Fine-tuning & Evals
Adversarially test a production LLM agent, document repeatable failures, and implement input/output guardrails and an eval suite that keeps them fixed.
What we built
Before their agent goes public, they need someone to break it — prompt injection, jailbreaks, data leaks — and then build the guardrails.
Benchmark Collective · Fine-tuning & Evals
Run a structured bake-off across models and prompts on the client's actual task set, then recommend a config with an accuracy-vs-cost tradeoff writeup.
What we built
They want data, not vibes: which model + prompt is most accurate and cheapest for their reconciliation agent's real tasks.
Groundtruth Labs · Fine-tuning & Evals
Build a pipeline that generates and curates training data for a domain fine-tune: synthesize examples, filter with quality checks, and produce a validated dataset.
What we built
They need a clean, deduped, high-quality dataset to fine-tune a small model — generated, filtered, and validated by an automated pipeline.
Draft House · Workflow automation
Build a copy-generation agent grounded in a brand style guide and product catalog, producing on-voice drafts for human review inside their CMS.
What we built
An agent that drafts campaign copy in the client's exact brand voice, pulling live product facts so nothing is invented.
Clearline Legal Ops · Workflow automation
Build a contract-analysis agent that extracts obligations, dates, and risky clauses from uploaded agreements and produces a reviewer-ready summary.
What we built
An agent that reads vendor contracts, extracts key terms, and flags clauses that break company policy — a first-pass reviewer that never tires.
Longform Media Co. · Workflow automation
Build a pipeline that takes long-form transcripts and generates show notes, timestamped highlights, and platform-specific social posts for review.
What we built
An agent that turns each episode transcript into show notes, clips, and social posts on brand — a content team in a box.
Kindred Helpdesk · Support agents
Build an agent-assist copilot embedded in their helpdesk that suggests grounded reply drafts and surfaces relevant context in real time.
What we built
A copilot beside human reps that drafts replies, pulls the right KB article, and suggests next steps — cutting handle time in half.
Lucid Systems · RAG & Search
Audit an existing RAG pipeline and lift answer quality: fix chunking, add reranking, tighten prompts, and prove the gain with before/after evals.
What we built
Their RAG app works 70% of the time — they need someone to close the gap with better chunking, reranking, and prompts.
Northgate Product Studio · Support agents
Build an in-product assistant that answers usage questions from the docs and executes common actions via the app's API, all inside a sidebar UI.
What we built
A copilot that answers 'how do I do X in this tool?' and can actually take the action for the user — support and productivity in one.
Steady Ground HR Tech · Workflow automation
Build an onboarding agent that answers new-hire questions from policy docs and orchestrates the setup checklist across their HR systems.
What we built
An agent that guides new hires through paperwork, answers policy questions, and files everything — so HR stops chasing forms.
Second Draft EdTech · Support agents
Build an adaptive tutor agent over their curriculum that answers questions, checks understanding, and adjusts explanations to each learner.
What we built
A tutoring agent that notices where a learner is stuck, explains it a different way, and adapts — grounded in the course material.
Meridian Retail Analytics · Data & Analytics
Build a conversational analytics agent over their data warehouse: translate questions to safe SQL, run it, and return summarized answers with the chart and the query it used.
What we built
A text-to-SQL agent that lets ops leads ask their warehouse plain-English questions and get charts back — no analyst in the loop.
Relay Support Co. · Support agents
Build a Tier-1 support agent integrated with their account systems that resolves common requests and hands off complex ones with full context.
What we built
The agent handles billing, plan changes, and outage FAQs end to end, and only routes true edge cases to human reps.
Resonant Voice AI · Voice agents
Build a conversational voice agent to replace a legacy IVR: natural intent handling, account lookups, and clean transfers to the right human team.
What we built
Replace the dreaded press-1-press-2 phone tree with a voice agent that understands what callers want and routes or resolves it.
Commons Collective · RAG & Search
Build a Q&A agent over their docs and forum history that gives cited answers in Discord and on-site, and defers to humans when unsure.
What we built
An agent that answers developer questions from docs, past threads, and code samples — cited, so the community trusts it.
Openworld Languages · Voice agents
Build a low-latency voice agent for spoken language practice that holds natural conversation, gives feedback, and tracks progress per learner.
What we built
A voice agent learners talk to for real conversation practice — it corrects gently, adapts to level, and never gets tired of repeating.
Rosewood Health AI · Fine-tuning & Evals
Stand up an automated evaluation pipeline for their patient-intake assistant: build a graded test set, wire LLM-as-judge scoring, and gate deploys on the results in CI.
What we built
Before they ship, they need to know the model is right — an eval suite that scores clinical accuracy and safety on every prompt change.
Haven Trust & Safety · Data & Analytics
Build a content-moderation pipeline that scores listings and messages against policy, auto-actions clear cases, and queues ambiguous ones with an explanation.
What we built
An LLM classifier that reads listings and messages, catches policy violations with nuance, and routes the gray-area cases to humans.
Verbatim Works · Workflow automation
Build a pipeline that ingests meeting transcripts, produces structured summaries and action items, and creates tasks in their PM system automatically.
What we built
An agent that turns raw call transcripts into clean summaries, decisions, and tasks pushed straight into the client's project tool.
Northstar Insurance Ops · Workflow automation
Build a claims-intake agent that validates submissions, extracts structured data from attachments, and routes by type and priority with a full audit trail.
What we built
An agent that reads incoming claims, checks completeness, extracts key fields, and routes each to the right queue — cutting days off the cycle.
Lumen Learning Studio · Fine-tuning & Evals
Build an evaluation agent that reviews generated course content against accuracy and clarity rubrics and blocks weak lessons before they ship.
What we built
An agent that reviews AI-generated lessons for accuracy and clarity, and scores whether they'll actually teach — a QA gate before publish.
Foundry DevTools · RAG & Search
Build a search agent over their scattered internal knowledge and expose it as an MCP server so their IDE and Slack bot can both query it. Freshness and access control matter.
What we built
Engineers ask questions in Slack and get answers pulled from Notion, GitHub, and Confluence — with an MCP server other tools can reuse.
Wayfinder Data · RAG & Search
Upgrade a keyword search to hybrid semantic search over a large content library and tune relevance with a real query eval set.
What we built
Their search returns near-misses; they need embeddings + reranking so results actually match what readers meant.
Meridian Global Support · Support agents
Build a multilingual support agent that detects language, answers from a shared KB, and keeps tone natural across locales, with human handoff per region.
What we built
One agent that supports customers in 12 languages, grounded in the same knowledge base, so the client skips staffing per-region teams.
Anchor & Co. · Workflow automation
Build a personal email agent that classifies, prioritizes, and drafts replies in the founder's voice, all held for one-tap approval.
What we built
An agent that sorts a founder's inbox, drafts replies in their voice, and flags what truly needs them — a chief-of-staff for email.
Foundry Knowledge · RAG & Search
Build a pipeline that clusters resolved support tickets, drafts help-center articles for the gaps, and feeds them back into a searchable knowledge base.
What we built
An agent that mines resolved tickets to draft the missing help-center articles — turning tribal knowledge into a searchable KB.
Tideline Freight · Workflow automation
Build a document-processing agent that ingests supplier invoices, extracts structured data, reconciles against purchase orders, and pushes clean records into their accounting system.
What we built
An agent that reads messy PDF invoices, extracts line items, matches them to POs, and flags mismatches — killing hours of manual entry.
Ledger & Proof · RAG & Search
Build a verification layer that cross-checks the app's generated claims against retrieved sources and flags or rewrites anything unsupported.
What we built
An agent that checks whether each generated claim is actually supported by its cited source — killing hallucinations before they ship.
Bright Harbor Brands · Data & Analytics
Build an analysis agent that ingests social mentions, classifies sentiment and topics, and produces a weekly narrative of what's moving and why.
What we built
An agent that reads thousands of mentions daily, clusters themes, and surfaces what's actually trending — not a keyword dashboard.
Cornerstone Data · Data & Analytics
Build an LLM-assisted data-quality agent that profiles tables, detects suspicious records, and produces human-readable issue reports for the data team.
What we built
An agent that scans datasets for anomalies an automated check would miss — plausible-but-wrong values — and explains what's off.
Northwind Legal Cloud · RAG & Search
Ship a retrieval-augmented support agent over their document corpus: chunk and embed the knowledge base, wire up hybrid search, and return cited answers inside their existing help widget.
What we built
The agent answers customer questions grounded in 4,000+ pages of contract law docs — with citations, so support reps trust every reply.
Willowbrook Health · Voice agents
Build an intake agent that gathers structured medical history via voice or chat, follows a clinical protocol, and hands a clean summary to the provider.
What we built
A voice-and-chat agent that collects patient history before appointments and structures it for the clinician — safely, with strict guardrails.
Routeworks · Workflow automation
Build a scheduling agent that handles inbound service requests, optimizes against technician availability and location, and confirms bookings automatically.
What we built
An agent that takes inbound job requests, checks technician availability, and books the slot — re-juggling in real time when plans change.
Harbor Goods Co. · Support agents
Build a support agent that auto-resolves order-status, returns, and FAQ tickets against their help center and order data, and hands complex cases to humans with context attached.
What we built
The agent reads every incoming ticket, resolves the easy 60%, and routes the rest with a drafted reply — so a 3-person team handles 10x volume.
Tideline Operations · Workflow automation
Build a Slack-native agent that exposes their common ops tasks as safe tool calls, so the team can trigger workflows and lookups by chatting.
What we built
A Slack agent the team messages to run routine ops — pull a report, kick off a workflow, check a status — without opening five tools.
Brightsmile Dental Partners · Voice agents
Build a natural-sounding voice agent that handles inbound scheduling end to end: understands intent, checks live availability, and writes the booking back to their practice-management system.
What we built
A phone agent that books, reschedules, and confirms appointments across 6 clinics — so front-desk staff stop drowning in calls.
Balance Ledger Co. · Data & Analytics
Build a reconciliation agent that ingests transactions, matches and categorizes them against the ledger, and surfaces exceptions with explanations.
What we built
An agent that matches transactions to records, categorizes them, and flags the messy exceptions a human should eyeball — clean books, faster.
Lumen B2B Growth · Workflow automation
Design and run a multi-step agent workflow that enriches new leads, scores fit, and drafts personalized first-touch emails for human approval. Ongoing retainer to tune and expand it.
What we built
An agent that researches inbound leads, drafts tailored outreach, and enriches the CRM — freeing the SDR team to actually close.