crenya
Admin sign-in required
or
🔒
Not an admin

Your account does not have admin access.
Add your email to ADMIN_EMAILS on the server to gain access.

Sign out
crenya
Sign out

Admin console

Overview, models, and runtime settings.

loading…
↻ refresh
Loading…
Users
Active paid
Auto-renew subs
MRR (est.)
Revenue · 30d
Revenue · all
Loading…
WhenUserPlanVia
Loading…
Live ops: tier-stats · provider-health · latency SLOs · queue

One cell per hour, oldest left. Grey = no traffic (not "healthy" — an idle hour and a clean hour are different facts). Window is capped at 72h because that is how long the hourly buckets live; anything longer would be invented history.

Loading…

Estimates, not invoices — incremented with the failover engine's per-request estimate, so a threshold crossing is meaningful even if the absolute $ is ±30%.

Loading…

Manually set a user's plan (no payment). Use free to revoke.

EmailPlanAuto-renewJoinedSet plan
Loading…
Page size

Real input/output tokens per model, with an approximate cost (provider list prices; cached input billed ~20%). Bars show input (blue) vs output (green) volume.

Loading…

Share of agent edit-tool calls that applied cleanly — detected server-side from edit results (the agent-quality north star). Compare the cheaper tool model vs the frontier one: if success % is comparable, keep the cheaper routing; if it drops, revert tool_call_model_cheap in Settings → Model routing.

Loading…

Which domain skill packs the agent actually loads (counted server-side from the injected skill block — no Langfuse). The thesis test: if the packs are rarely invoked, the domain-depth bet isn't landing yet.

Loading…
Live in ~30 s — changes apply within the config cache TTL after saving.

Estimated USD. These exist so an outage that re-routes the fleet onto an expensive fallback can be capped without a deploy — that is why they are editable here. The ceiling drops degraded fallbacks for free/trial only; paying tiers are never spend-blocked and no request is ever refused.

One-time setup: creates a monthly Razorpay plan per paid tier so auto-renew checkout works. Idempotent — safe to click again.

Live promo codes — add, edit, or remove instantly (no redeploy). percent = % off checkout; free_days = grants a plan free for N days with no payment. Removing a code stops it immediately; redemption counts are preserved.
CodeTypeValue PlanExpiresUsed / Max
Loading…
Measured model quality — the benchmark engine scores each tier/model on the coding-eval and ranks by pass@1. This is the source of truth for routing — measured, not vendor claims. Runs are stored versioned (never overwritten).
Loading…
Advisory only. Nothing switches automatically — the engine ranks by pass@1, breaks ties on $/solve, and you flip the slot yourself in Settings → Model routing. Benchmark scores are compared against live production signals (TTFT p95, edit-success) because the offline eval is blind to latency.
Loading…
#Targetpass@1PassedWall s
Loading…
TargetOverallEasyMediumHard
WhenLabelWinnerpass@1Targets
Loading…
Did anything break? — user-facing failures, so a silent breakage shows up here instead of waiting for a customer to report it. ide = reported by the app (the class the server can't see); server = what the backend knows (empty_response_healed = a user almost saw an error; _unhealed = they did).
TOTAL ERRORS
FROM APP (IDE)
FROM BACKEND
KindCount
Loading…
ModelKindCount
VersionErrors
Ops error log — provider degraded/failed, reliability, and billing alerts land here instead of email. Each issue is deduped to once per hour, newest first. Email is off by default — flip ops_email_alerts in Settings to also email during a critical window.
Loading…
What actually serves a request. These are the live values the router reads, not the defaults baked into the deploy. An A/B rule sends a slice of one tier to a challenger model — it was the only routing control with no page at all, so stopping one meant a database or a script.
Loading…
Sets model_ab to {}. Takes effect on the next request.
Edit these on the Settings tab. They are read on every request, so a change is live without a deploy.
Voice is billed per audio second and vision per image, so each has its own daily ceiling per plan.
Did users LIKE it? — every other board measures machinery (edit-success = the edit applied; the eval = we pass tests). This is the only one that says the human was happy. A 👎 is also a real failure case worth turning into an eval task.
SATISFACTION
👍 UP
👎 DOWN
BY TIER
Loading…
VersionRatings