AiMedley
About AiMedley

Build a team of AI experts.
Choose their brains.

AiMedley is where personas outlive vendors. Create persistent AI workers — give each a name, mission and personality — then pick which of 12 brains across 8 providers thinks for them today. Swap models tomorrow without losing anything.

12
Brains
8
Providers
0
Lock-in
GPT-5.1OpenAI
GPT-4o miniOpenAI
Sonnet 4.6Anthropic
Haiku 4.5Anthropic
Gemini 2.5 ProGoogle
Flash-liteGoogle
DeepSeek V3DeepSeek
Qwen 2.5 72BAlibaba
Kimi K2Moonshot
Llama 3.3 70BMeta
Mistral LargeMistral
Auto BrainRouter
GPT-5.1OpenAI
GPT-4o miniOpenAI
Sonnet 4.6Anthropic
Haiku 4.5Anthropic
Gemini 2.5 ProGoogle
Flash-liteGoogle
DeepSeek V3DeepSeek
Qwen 2.5 72BAlibaba
Kimi K2Moonshot
Llama 3.3 70BMeta
Mistral LargeMistral
Auto BrainRouter
The flagship
Debate mode

Put your workers in the ring.

Same prompt. 2–5 workers. Each argues a different stance. A judge worker weighs the arguments and delivers a structured verdict — live, in parallel, in under a minute.

  • ✓Parallel SSE streaming — see all sides at once
  • ✓Structured verdict: winner, why, where each side fell short
  • ✓Every debate auto-saved & one-click shareable
  • ✓Typical cost: ~$0.01 per debate
Topic
Should companies mandate quarterly in-person offsites?
Round 1 · Live
👩‍💼NinaFor
Async slays daily standups, but relationships get built in the messy edges…
🧑‍💻RioAgainst
Every mandated offsite is a productivity tax on introverts who ship best solo…
Verdict by Judge Marla
Winner: Nina (narrow). Rio's productivity math holds, but Nina's "trust IOU" model ships better long-run.
Live SSE · $0.008 · 41s
Why we built it

Four beliefs behind AiMedley.

Persona outlives vendor.

Every LLM release resets your setup elsewhere. Not here. Your worker keeps their name, mission, memory, and history no matter which brain you swap in.

Multi-brain by default.

12 models across 8 providers behind one unified interface. Bring your own OpenAI / Anthropic / Google keys to bypass the middleman and cut cost 30-70%.

The debate is the demo.

Give the same prompt to 5 workers with different stances. A judge weighs the arguments in a live streamed verdict. Every debate is shareable in one click.

Your workforce, your keys.

Workers, conversations, debates, and API keys are scoped to your account. Direct provider keys are encrypted at rest. No training on your prompts. No ads.

The catalog

Twelve brains. Eight providers. One interface.

Every brain behaves the same way to your workers — swap them like keyboard shortcuts. Bring your own keys where you can to cut cost and latency.

OpenAI
GPT-5.1Reasoning
GPT-4o miniFast & cheap
Anthropic
Sonnet 4.6Writing
Haiku 4.5Speed
Google
Gemini 2.5 ProLong context
Flash-liteCheapest
DeepSeek
DeepSeek V3Math & code
Alibaba
Qwen 2.5 72BMultilingual
Moonshot
Kimi K2Storyteller
Meta
Llama 3.3 70BOpen weights
Mistral
Mistral LargeEuropean
Router
Auto BrainAuto-pick
What's shipped

Nine tools. Zero lock-in.

Persistent workers

Create AI colleagues with names, missions, memories, and expertise. Chat with each independently.

Brain swap

Change which LLM powers a worker anytime. Persona, memory, and history are preserved.

Debate mode

2–5 workers argue their stances. Judge worker delivers a structured verdict. Live SSE streaming.

Teams

Group workers under a leader with shared instructions and memory. Broadcast prompts to the whole team.

Templates gallery

Save any worker as a template. Clone the community's best setups in one click. Publish yours.

Public share links

Templates, conversations, and debates become no-auth URLs. Snapshot-based — no live LLM leak.

Compare mode

Fan the same prompt out to up to 4 brains simultaneously and watch them stream side-by-side.

Direct provider keys

Store your OpenAI/Anthropic/Google keys once (encrypted). Every worker on those brains bypasses OpenRouter.

Cost budgets

Daily and monthly caps per worker or globally. Live spend counter. Warn at 80%, block at 100%.

Under the hood

How it works.

1
Register → workforce seeded in 30 seconds.

The moment you sign up we spin up 4 demo workers + 1 team so you can run a debate immediately. No setup, no config, no prompt engineering.

2
One prompt, many brains.

Behind every worker is a unified `chat_stream(brain_id, msgs)` interface. Switch a worker from GPT-5.1 to Claude Sonnet 4.6 to Gemini 2.5 Pro with a dropdown.

3
Bring your own keys (optional).

Save your OpenAI / Anthropic / Google keys once. Workers on those brains bypass OpenRouter and use the vendor natively — cheaper, faster, no middleman margin.

4
Debates run in parallel, judged sequentially.

Each round fans debaters out in parallel via server-side asyncio queues. Once every debater finishes, the next round starts. The judge sees the full transcript and streams a structured verdict.

5
Share = snapshot, not live query.

When you share a template, chat, or debate we freeze the payload. Public viewers never re-trigger LLM calls or leak live worker state. You can safely edit the source after.

Stack

Boring, dependable pieces.

FastAPI
Async Python backend
MongoDB
Workforce state
OpenRouter
Default brain gateway
Direct SDKs
OpenAI · Anthropic · Google
React 19
Frontend
Tailwind
Styling
shadcn/ui
Primitives
SSE
Realtime streaming
For investors & partners
See the 14-slide pitch.
Vision, market, competition, roadmap. Keyboard-navigable. 3 minutes end-to-end.
Open the deck

Ready to hire?

Spin up a 4-worker team in 30 seconds. Run your first debate for a penny. No card, no lock-in.

Create my workforce