Skip to content
@evolvent-ai

Evolvent AI

Building persistent agent infrastructure for infinite self-evolving intelligence.

Popular repositories Loading

  1. ClawMark ClawMark Public

    🦞 ClawMark: A Living-World Benchmark for Multi-Day, Multimodal Coworker Agents

    Python 124 11

  2. RSIBench-Data RSIBench-Data Public

    Synthetic data generation, post-training, and E2B benchmark evaluation infrastructure.

    Shell 123 8

  3. Terrarium Terrarium Public

    Terrarium: Multi-turn data engine for evaluating and optimizing LLM agents in living environments.

    Python 59 2

  4. VibeLifeBench VibeLifeBench Public

    🗓️ The hardest life-admin benchmark for agents — lawsuits, escrow shortfalls, apartment hunts, exams. 20 long-horizon tasks × 20–30 stages across 10 domains and 21 services, scored by 1247 atomic c…

    PLpgSQL 13 1

  5. BenchRouter BenchRouter Public

    BenchRouter - Benchmark routing and evaluation framework

    Python 12

  6. Authbench Authbench Public

    Python 10 1

Repositories

Showing 10 of 11 repositories

Top languages

Loading…

Most used topics

Loading…