Stop surprise AI bills. Cap spend per project in 2 minutes.

The AI spend firewall. Route every model call through one OpenAI-compatible gateway — your keys, hard budget caps, alerts at 50/80/100%, and usage history you can trust.

Start Free →
LayerFlow Gateway
Hard cap
Monthly budget
$120.00 · $86.40 used (72%)
Alerts at 50/80/100%50% ✓ sent

You got emailed at 50%. Calls keep flowing until 100%, then they hard-block — no surprise bills, ever.

Attributed per projectx-lf-project · $0 wasted
The problem

AI is powerful.
Your workflow isn't.

Every time you hit a rate limit at 11 PM or switch between ChatGPT, Claude, and Gemini, you lose context, re-explain the project from scratch, and pay twice for the same tokens. LayerFlow preserves your project memory, auto-compresses context, and finishes the thought in under a second.

100%
BYOK & zero markup
$0
wasted on repeated tokens
★★★★★
browser + CLI sync
Try LayerFlow Free →
How it works

Three steps. From plain English to shipped code.

> write plain prompt...
STEP 01

Write in plain English

No prompt engineering ritual or complex setup. Type a simple sentence or paste a broken AI chat you got stuck on. LayerFlow extracts the real objective.

Score: 98/100
→
-8,240 tokens
STEP 02

Auto-improve & cut context

LayerFlow refines your request into a precise prompt with strict constraints, tests, and schema, while trimming away token bloat to keep your costs near zero.

✓ Browser
✓ CLI Sync
✓
STEP 03

Multi-agent run & sync

Implementation, code review, and test agents run in parallel across your choice of models. Every finished thought generates a Continue Pack ready for any AI.

Browser + Terminal

A terminal that
runs like an engineering team.

Install the lf CLI once. It auto-improves your prompt, trims unused repository context, runs parallel agents with approvals, and keeps one shared memory between your browser and command line.

01
11 PM limit reached
→

Rescue dead conversations

Paste any frozen or rate-limited AI chat. LayerFlow compresses it, extracts the architecture decisions, and generates a Continue Pack ready to drop into another model.

Never re-explain your project at midnight
02
Claude 3.7
DeepSeek R1

Multi-model routing & BYOK

Route reasoning tasks to DeepSeek R1, code generation to Claude 3.7, and speed tests to Groq. Use your own keys with zero markup and hard budget caps.

Support for OpenAI, Anthropic, Gemini & DeepSeek
03
lf
+
run
⌘
+
K

Unified web + terminal sync

One workspace. Start a coding task in the CLI while traveling, inspect the agent diffs in the web UI, and approve terminal file modifications with one tap.

Instant bi-directional state synchronization
Build your own

Build your own agent
and just chat.

LayerFlow puts you in control. Bring your own API keys, cap your budget, wire up tools, and run an agent that does the work — or skip all of it and simply chat. Same workspace, two ways in.

01
lf
+
agent

Build your own agent

For coders: attach your own model keys, add tools and approvals, set hard budget caps and request limits. Your agent plans, writes, and ships — with every cost line visible.

Bring your keys · BYOK · zero markup
02
⌘
+
K

Or just chat

No code, no setup. Open a chat, type plain English, and get real answers the moment you hit “Start”. Perfect for everyone who wants the outcome, not the terminal.

Chat first · no terminal needed
03
∞

Runs fully local

Use Ollama or LM Studio on your own machine — no account, no cloud, no API key required. The terminal, models, and your conversations all live where you do.

100% local · like opencode
Build yours — Start Coding →

Start free with your own keys or a local model. No credit card, no lock-in.

Agent at work

An agent for the job hunt.
apply, pitch, repeat.

Build your own job-hunt agent. It scans openings, tailors your resume and cover letter, applies while you sleep, and hunts freelancing clients with pitches that sound like you — all in the background.

01
lf
+
job

Auto-apply to jobs

Your agent watches job boards and companies for roles that match your skills, ranks them by fit, rewrites your resume and cover letter per position, and submits the application — no copy-pasting.

Cover letters written for every role
02
lf
+
client

Find freelancing clients

Point it at Upwork, Fiverr, or your own target list. It matches you to briefs, writes a pitch in your tone, sends the first message, replies to follow-ups, and nudges leads before they go cold.

Pitches that sound like you
03
⌘
+
A

Runs in the background

The grunt work never sleeps. Your agent tracks replies and deadlines, surfaces next steps, and pings you only when a human decision is needed — one tap to steer it.

You approve · the agent does the rest
Build your job agent →

Start free with your own keys or a local model. No credit card, no lock-in.

What makes us different

Others chat in silos.
LayerFlow finishes the job.

Standard AI chat tools are isolated web tabs. They lose context the moment you switch models, leave you stranded at midnight limits, and charge you for repetitive tokens.

Standard AI Chatbots

✕Forget everything when you switch models or tabs
✕Hit a limit at 11 PM and lock you out with half-written code
✕No local terminal sync, forcing manual copy-pasting of files
✕Repeatedly send full conversation history, wasting token dollars
vs

LayerFlow

✓One persistent memory shared across Claude, GPT-4o, and DeepSeek
✓Instantly creates Continue Packs to rescue dead or limited chats
✓Native lf CLI with approval prompts directly in your shell
✓Auto-cuts unused context & enforces hard spending budget limits
LayerFlow is not just another wrapper. It is the intelligence and memory your AI workflow was missing.
Engineered in the open

No lock-in.
No token markups. No BS.

LayerFlow was created by developers who got tired of AI context rot, midnight rate limits, and paying twice for the same prompt history. Bring your own API keys, run locally, and control every single penny.

< 0.8s
prompt refinement & context cut
$0.00 markup
direct provider billing with BYOK
0 lost chats
instant rescue into Continue Packs
Pricing

Transparent pricing. Zero token markup.

Start free with your own API keys. Upgrade when you need unlimited gateway requests, a shared key vault, and team-wide budget control.

Developer

Free forever
$0no card needed
  • 100% BYOK (OpenAI, Claude, DeepSeek, Groq)
  • Gateway capped at 20 messages/day — proof it works
  • Hard budget caps + 50/80/100% alerts
  • Usage history
Start Free
Most popular

Pro

For engineers & builders
$19per month
  • Everything in Developer
  • Unlimited gateway requests (no daily cap)
  • BYOK vault — provider keys + custom base URLs
  • Per-project spend + priority support
Get Pro Access

Team

For engineering squads
$49per month
  • Everything in Pro
  • Team-wide budget caps & shared key vault
  • Centralized BYOK vault & audit log
  • Priority support & private Discord
Upgrade Team
Prices in USD. All plans support 100% Bring-Your-Own-Key (BYOK) with 0% token margin. Questions? Read the key security policy.
FAQ

Everything you would ask before switching your AI workflow.

What is LayerFlow and how is it different from standard AI chat apps?+
LayerFlow is an AI spend firewall and gateway. One OpenAI-compatible endpoint sits in front of every major model — your keys, hard budget caps, alerts at 50/80/100%, and usage history. Rather than locking you into one vendor, it lets you route calls to Claude, ChatGPT, Gemini, and DeepSeek while keeping every dollar accounted for.
How do hard budget caps work?+
You set a monthly cap per workspace. Every request counts against it, and at 50%, 80%, and 100% you get an email. At 100%, new requests hard-block — no surprise bills, no exceptions.
Can I bring my own API keys (BYOK)?+
Yes! Plug in your own keys for Anthropic, OpenAI, Google Gemini, DeepSeek, Groq, and more. LayerFlow charges 0% token markup — you pay provider wholesale rates directly, and we add caps, alerts, and usage history on top.
How does per-project spend attribution work?+
Send one header — x-lf-project: client-acme — on any request and the cost lands under that project in your usage history. No spreadsheets, no guesswork.
What exactly is usage history?+
Every gateway completion records the model, input/output tokens, project tag, and the exact dollar cost to four decimals. View it in the dashboard, or via `lf cost` in the terminal.
What AI models are supported?+
All major frontier models: Claude Sonnet / Opus, OpenAI GPT-4o / o3-mini, Google Gemini Flash / Pro, DeepSeek R1 / V3, Groq (Llama 3.3 70B), and more via a single OpenAI-compatible API.
Can I use the CLI without a LayerFlow account?+
Yes. Set a provider key (`lf config key openai sk-...`) or an env var like OPENAI_API_KEY and lf chats directly to that provider at your configured base URL. `lf login` is only needed for gateway budgets.
Is my code private and secure?+
Yes. With BYOK, requests go straight to provider endpoints through your own key. Direct-mode requests are never stored. LayerFlow never trains on your code or prompts, and provider keys in the vault are encrypted.

Stay ahead of frontier models.

Get an email when new model integrations, CLI workflows, and prompt templates drop. Zero noise.