Caspian Labs

The search engine for opportunities.

Create better ideas.
Improve what you’re building.

Explore CASP
A NEW WAY TO EXPLORE

Your next idea could
be a much better one.

Bring a goal, an idea, or something you’ve built. CASP handles the research and idea development. Guide it at any point, or let it develop the result for you.

MEET CASP

Find what’s worth pursuing.

Start with a question. Discover a new direction through deep research.

Find a new way to keep city homes cool.

Deep researchRun example

Interactive example · real sources · illustrative opportunities

CASP / RESEARCH EXAMPLE

A cooling business hidden in a noise problem.

What it changes

What remains to be tested

Read original source
01

Find an unexpected connection.

Explore across markets, technologies, and disciplines to uncover opportunities beyond the reach of standard AI research.

02

Develop it until it holds up.

Work out how the idea would function. Investigate weak assumptions and use what’s learned to improve it.

03

Leave with work you can use.

A developed concept, product improvement, or plan, with a clear next action.

BUILT ON DEEP RESEARCH

Better ideas need
a wider view.

On the 45-task WANDR comparison, CASP outperformed every published production research system.

WANDR 45-task leaderboard

  1. CASP
    0.5833
  2. Perplexity
    0.447
  3. Anthropic
    0.262
  4. OpenAI
    0.153
  5. Exa
    0.111
  6. Parallel
    0.080
  7. Gemini
    0.074
WANDR BENCHMARK

Compared with the published leader,
Perplexity Search as Code.

+30.5%Research quality & coverageSoft F1
+120%Fully evidenced resultsHard F1
ONE ENGINE. MORE POSSIBILITIES.

Wherever you do
your best thinking.

We’re building CASP for people, teams,
and the tools they already use.

In development. Join us early.
THERE’S MORE AHEAD

What could you
create next?

EARLY ACCESS

Bring your
next possibility.

Tell us what you’d like to create or improve. We’ll review your goal and get in touch about early access.

CASP’s commercial experience is in development. We’ll use these details to reply about your request. No account or payment required.
THE RESEARCH FOUNDATION

The full comparison.

CASP’s recorded results and every system configuration in WANDR’s published 45-task table.

CASP EVALUATION

45 recorded task results. Per-task scores, receipt hashes, and evaluation notes.

Results & methodology (opens in a new tab)Task-level scores (opens in a new tab)
PUBLISHED COMPARISON

The source for the 19 published system configurations below.

WANDR Table 6 (opens in a new tab)
Ordered by Soft F1Higher is better ↗

Swipe across for scores

CASP official evaluator results and all 19 published configurations from WANDR Table 6.
SystemSettingCompletedSoft F1Hard F1
CASPCustom45 / 450.58330.4926
Perplexity Search as Codexhigh44 / 450.4470.224
Perplexity Search as Codehigh45 / 450.3970.156
Perplexity Search as Codemedium45 / 450.2950.149
Anthropic managed agenthigh45 / 450.2620.099
OpenAI web agenthigh45 / 450.1530.073
OpenAI web agentxhigh45 / 450.1270.060
Perplexity Search as Codelow45 / 450.1210.053
Exa Agentxhigh45 / 450.1110.036
OpenAI web agentmedium45 / 450.0910.053
Parallel Tasksultra8x45 / 450.0800.035
Gemini Deep Researchmax45 / 450.0740.028
Exa Agenthigh45 / 450.0730.029
Parallel Tasksultra4x45 / 450.0670.025
Parallel Tasksultra2x45 / 450.0630.026
Gemini Deep Researchspeed45 / 450.0480.020
OpenAI web agentlow45 / 450.0380.024
Exa Agentmedium45 / 450.0130.004
Exa Agentlow45 / 450.0060.003
Parallel Tasksultra40 / 450.0050.002

Soft F1 measures research quality and coverage. Hard F1 credits results that meet their full evidence requirements. CASP reports the mean of its best preserved official evaluator results across 45 tasks. Reference scores come from WANDR Table 6.

Built by Atrin Farnamfar.

THE EXPERIENCE WE’RE BUILDING

Planned product experience. Availability is not yet announced.