Evidence, not vibes
Source hashes, span-level excerpts, claim graphs, freshness tracking, and stale-claim quarantine keep every conclusion tied to what was actually observed.
AUTONOMOUS RESEARCH INFRASTRUCTURE
Evidra is a research harness for long-running investigation and experimentation. It turns questions into evidence, evidence into experiments, and experiments into decisions that hold up.
Most agent tooling is built around a conversation or a single code change. Evidra is built around the work that happens after the first promising answer.
Reasoning is only one part of research. Evidra owns the deterministic layer that makes autonomous work inspectable, reproducible, and harder to fool.
Source hashes, span-level excerpts, claim graphs, freshness tracking, and stale-claim quarantine keep every conclusion tied to what was actually observed.
Focus-aware lanes explore different formulations, exchange compact evidence boards, and spend peer review where disagreement is useful.
Falsification agendas and negative evidence stop the harness from rediscovering the same dead ends.
Checkpoints, leases, heartbeats, crash recovery, budgets, and resume semantics keep a campaign alive across machines and sessions.
Evaluator-integrity protection, code-health checks, best-so-far ratchets, paired statistics, replication requirements, and verified-state gates separate progress from self-deception.
Declare a metric, artifact, proof, behavior, or system property. The harness adapts to the objective—not the other way around.
One durable loop from open question to validated result. Every phase has an internal goal, a budget, and a reason to stop.
Start in the terminal. Give it a repository, a challenge, or a research question. Evidra coordinates the rest.
Evidra brings modern agentic research techniques into one accountable runtime. Each mechanism is bounded, observable, and connected to measured outcomes.
Independent islands, family diversity, migration, survivor selection, and evaluator-backed crossover proposals explore more than one path.
Failed hypotheses become negative evidence. The next cycle must change the mechanism, test, or execution route.
Recorded discovery trees can be replayed under alternative policies to compare branch order, cost, parallelism, and Pareto outcomes.
Hashed sources, span-level excerpts, claim graphs, freshness checks, and quarantine keep conclusions traceable.
UCB portfolios, information gain, uncertainty, diversity, budget, and evidence pressure determine where the next unit of work goes.
Independent critics, paired statistics, multi-seed checks, sequential-look correction, and replication gates protect against false wins.
Changes to Evidra itself are treated as experiments with component snapshots, predictions, matched tasks, and measured retain or revert decisions.
Timeouts, rate limits, worker crashes, OOMs, and invalid outputs trigger classified, route-changing recovery instead of blind retries.
Evidra is open source and runs from your terminal. Install it globally, authenticate your provider, and start a campaign in minutes.
Read the user guide ↗$ curl -fsSL https://raw.githubusercontent.com/
StarAtNyte/evidra/master/install.sh | shevidra/login codex/research start