Revalvo icon

Revalvo

Revalvo is a browser-based, local-first LLM eval workbench for running prompts across multiple models, scoring results, versioning prompts, and batch-testing datasets. It is designed for users who want to compare outputs and keep prompt work in the browser without creating an account.

Revalvo

Overview

Revalvo is a local-first LLM eval workbench for writing, running, versioning, and evaluating prompts across multiple models. The homepage positions it as a place to run the same prompt against every model at once, score results, and keep a traceable record of what was shipped.

It runs in the browser with no account and no hosted database. The site says you can connect providers with your own keys, use local runtimes such as Ollama or LM Studio, or work with custom endpoints, then move from single prompt tests to dataset-based batch evaluation and reporting.

Core features

Parallel multi-model runs

Run the same prompt against multiple models in parallel and compare the outputs side by side in synced columns, with per-model timing and cost shown on the run.

Scoring and reporting

Score responses in real time and review report outputs that surface pass rate, model ranking, per-case pass/fail, and debug details.

Prompt version history

Save a run as an immutable snapshot, compare versions with model config diff and line-level word diff, and roll back to any earlier version.

Dataset evaluation workflows

Create datasets by hand, upload CSV files, start blank, or generate rows from a description; batch-run saved versions across the dataset.

Provider connection options

Connect providers through API keys, OAuth, localhost runtimes, or custom endpoints, with workspace separation per provider.

Export and handoff options

Export saved versions or batch reports as Python, TypeScript, cURL, Node.js, Go, agent prompts, HTML, Markdown, JSON, or CSV.

Practical use cases

  • Choose a model for a prompt

    Compare one prompt across several models at once, review the outputs side by side, and see which model scored best for the chosen criteria.

  • Iterate on prompt drafts

    Treat prompts like code by saving immutable versions, comparing configuration and word-level changes, and rolling back if a later draft performs worse.

  • Test against a dataset before release

    Build a dataset, run a saved prompt version across it, and use the report to check pass rate, latency, cost, and failures before shipping.

  • Evaluate prompts in a local-first setup

    Connect a provider workspace with your own key or local runtime, then work in a browser without creating an account or using a hosted database.

  • Hand off results to other tools

    Export prompt versions or batch results into code snippets, agent prompts, or common file formats for handoff to a coding agent or teammate.

Pros and Cons

Pros

  • Local-first browser workflow with no account required.
  • Parallel runs make it easy to compare the same prompt across multiple models.
  • Version snapshots, diffs, and rollback are built into the prompt workflow.
  • Batch evaluation supports datasets, scoring, and structured reports.
  • The privacy policy says prompts, versions, datasets, eval configs, results, and keys stay in browser storage by default.

Cons

  • The pricing page is missing; /pricing returns a 404, so published plan details are not available on the site.
  • Some provider types are proxied through Revalvo's Worker, while others connect directly; the privacy policy distinguishes these flows, so data handling depends on the provider you choose.

FAQ

What is Revalvo used for?

Revalvo is a browser-based prompt engineering and LLM evaluation workbench. It lets you run prompts across multiple models, review scored outputs, version prompts, and export or roll back saved versions.

Which model providers does Revalvo support?

The site says you can connect providers by pasting a key, signing in with OAuth, or pointing to localhost. Supported options shown include OpenRouter, Ofox.AI, Vercel AI Gateway, Groq, OpenAI, Anthropic, Ollama, LM Studio, and custom OpenAI-compatible endpoints.

Does Revalvo store my prompts and keys on its servers?

Revalvo is designed to work in the browser with no account and local storage for prompts, versions, datasets, eval configs, results, and API keys. The privacy policy says direct providers can be used without Revalvo servers seeing your content, while some providers are proxied through a worker in memory only.

Can I run dataset-based evaluations?

The homepage shows a workflow for single runs in the playground and batch evaluation against a full dataset. Reports include pass rate, latency, cost, model ranking, per-case pass/fail, and export options in HTML, Markdown, JSON, and CSV.

Is there published pricing information?

No pricing page is available at /pricing; that URL returns a 404 on the site. The terms page states that Revalvo is free to use with no uptime SLA, but it does not describe paid tiers or plan details.

Quick Facts

Category
Developer Tool
Platform
Browser-based
Primary workflow
Prompt evaluation, versioning, and batch scoring
Data storage
Local browser storage by default
Pricing page
Not published; /pricing returns 404
Source domain
revalvo.com

Revalvo 대안

Edgee icon

Edgee

Edgee is an AI gateway for coding agents and LLM-powered apps. It compresses token traffic, routes requests across models, and provides observability and team controls to help reduce cost and keep sessions running.

ByteAsk icon

ByteAsk

ByteAsk is a terminal-first AI coding agent for C and C++ that edits repositories and verifies changes with the real compiler, debugger, sanitizers, and tests before showing a diff. It offers a free tier plus paid plans, with editor connectors and zero-retention handling described in the source.

Prompty Town icon

Prompty Town

Prompty Town is a web product that turns a link into a building in a small internet city. It appears to let users buy a tile, add a prompt, and publish the result alongside other buildings.

Manta AI icon

Manta AI

Manta AI is an autonomous web app testing tool for teams that want to map application behavior, catch regressions, and generate tests without writing scripts or maintaining selectors. It works from a URL and supports plain-English test flows, run results with screenshots, and scheduled or deployment-triggered checks.

Creativly icon

Creativly

Creativly is a web-based AI creative studio for generating visual concepts, mockups, and stylized images from short inputs. It is aimed at designers, creators, and entrepreneurs who want fast visual ideation without writing long prompts.

AakarDev AI icon

AakarDev AI

AakarDev AI helps teams manage AI provider access, project-level setups, logs, and analytics from one dashboard. It supports BYOK workflows and lists providers including OpenAI, Google Gemini, Anthropic, Groq, Mistral AI, and Perplexity AI.

Revalvo - AI Tool, Features, Use Cases & Alternatives | UStack