๐Ÿ“„๐ŸŽ‰ Our paper is live on bioRxiv!NEW

The OpenClaw Ecosystem for Scientific Research

A curated dataset and discovery platform for OpenClaw AI agent projects โ€” spanning bioinformatics, biomedicine, drug discovery, and multi-agent learning. From lightweight assistants to self-evolving research systems.

OpenClaw Ecosystem Projects

From the full-featured core platform to ultra-lightweight edge agents, from self-evolving research systems to multi-omics pipelines.

Core Platform & Lightweight Variants

OpenClaw

OpenClaw

383.7K

The core platform. Personal AI assistant on your own devices. 25+ messaging channels (WhatsApp, Telegram, Slack, Discord, Signal, iMessage, Teams, Matrix, Feishu, LINE, WeChat, QQ). Voice on macOS/iOS/Android. Live Canvas rendering. Gateway control plane. 358K+ stars.

ai-assistantmulti-platformvoicecanvas
TypeScript
Hermes Agent

Hermes Agent

218.3K

The self-improving AI agent by Nous Research. Built-in learning loop โ€” creates skills from experience, improves them during use, nudges itself to persist knowledge, cross-session recall via FTS5 search. 200+ models via OpenRouter. 6 terminal backends (local, Docker, SSH, Daytona, Singularity, Modal). Telegram/Discord/Slack/WhatsApp/Signal. Cron scheduler. Subagent delegation. Research-ready with batch trajectory generation and Atropos RL environments.

Nous Researchself-improvinglocal-modelsper-model-parserlearning-loop
Python
Pi

Pi

74.6K

Independent AI agent harness mono repo from earendil-works (Mario Zechner / badlogic). Includes pi-ai (unified multi-provider LLM API: OpenAI, Anthropic, Google), pi-agent-core (runtime with tool calling + state), pi-coding-agent (interactive CLI), and pi-tui (terminal UI with differential rendering). Companion pi-chat for Slack/chat workflows. Strong supply-chain hardening (pinned deps, min-release-age). Foundation under feynman โ†’ Darwin.

agent-harnesscoding-agentmulti-provider-LLMmonorepo
TypeScript
NanoBot

NanoBot

46K

Ultra-lightweight personal AI agent โ€” 99% fewer lines of code than OpenClaw. v0.1.5: Dream skill discovery, mid-turn follow-up injection, WebSocket channel, context auto-compact, notebook editing, multiple MCP servers. Supports Feishu streaming, QQ, WeCom, Kagi web search. From HKUDS (ClawTeam team). 39K+ stars.

lightweightmulti-modelMCPsub-agents
Python
ZeroClaw

ZeroClaw

32.3K

Fast, small, and fully autonomous AI assistant infrastructure. Deploy anywhere, swap anything. Written in Rust with extreme performance focus โ€” the infrastructure layer for self-hosted AI agents.

infrastructureautonomousdeploy-anywhereRust
Rust
NanoClaw

NanoClaw

30.3K

Lightweight alternative to OpenClaw that runs in containers for security. Connects to WhatsApp, Telegram, Slack, Discord, Gmail. Built on Anthropic Agents SDK with memory and scheduled jobs.

containersecureAnthropic-SDKlightweight
TypeScript
PicoClaw

PicoClaw

29.7K

Tiny, fast, and deployable anywhere. Ultra-efficient AI assistant written in Go by Sipeed. Designed for edge devices and resource-constrained environments. Inspired by NanoBot.

edgeIoTRISC-Vtiny
Go
NemoClaw

NemoClaw

21.9K

NVIDIA official reference stack for running OpenClaw securely inside NVIDIA OpenShell sandbox. Managed inference with Nemotron open-source models. Essential for labs running autonomous agents on GPU servers.

NVIDIAsandboxsecurityGPUNemotron
JavaScript
OpenHarness

OpenHarness

15K

Open Agent Harness from HKUDS (NanoBot team). Runtime environment and execution framework for AI agents. 615 stars in 1 day.

HKUDSharnessruntimeagent-infrastructure
Python
IronClaw

IronClaw

12.5K

Rust implementation of OpenClaw focused on privacy and security. Memory-safe architecture with Rust's ownership model, built by NEAR AI. Ideal for sensitive research environments requiring data isolation and high performance.

Rustprivacysecurityhigh-performance
Rust
NullClaw

NullClaw

7.8K

Fastest, smallest, fully autonomous AI assistant infrastructure written in Zig. Part of the Null ecosystem (NullHub UI, NullWatch observability, NullTickets task tracking). Extreme performance with minimal resource footprint.

Zigfastestsmallestautonomous
Zig
autoagent

autoagent

4.5K

Autonomous harness engineering. Self-configuring agent runtime that adapts to your workflow. 560 stars in 24 hours.

harnessautonomousruntimeself-configuring
Python
FastClaw

FastClaw

1.3K

Faster and lighter OpenClaw alternative written in Go. Focused on speed, simplicity, and low resource usage. Quick to deploy for researchers who want minimal setup overhead.

Gofastlightweightalternative
Go
EdgeClaw

EdgeClaw

1.2K

Edge-Cloud Collaborative AI Agent built on OpenClaw โ€” bringing the Claude Code experience to open source. v2.0 features ClawXMemory (multi-layered structured long-term memory with proactive reasoning) and ClawXRouter (LLM-as-Judge cost-saving router achieving 58% savings). Three-tier privacy (S1 Passthrough / S2 Desensitization / S3 Local), visual Dashboard, zero-config setup. By THUNLP, RUC, OpenBMB.

edge-cloudprivacymemorycost-savingTsinghuaOpenBMB
TypeScript
MicroClaw

MicroClaw

729

Agentic AI assistant built with Rust. Runs on <1MB RAM targeting $2-5 MCU hardware like ESP32. The smallest member of the Claw family for sensor-level embedded AI.

embeddedMCUESP32rust
Rust

Team & Orchestration

Multica

Multica

41.4K

Open-source managed agents platform. Turn coding agents into real teammates โ€” assign tasks, track progress, compound skills. Agents show up on boards, post comments, report blockers autonomously. Works with Claude Code, Codex, OpenClaw. Self-hosted or cloud (multica.ai).

managed-agentstask-managementmulti-agentself-hosted
TypeScript

Edict

16.3K

Multi-agent orchestration inspired by China's Imperial Three Departments and Six Ministries system. 9 specialized AI agents with real-time dashboard, model config, and full audit trails. You are the Emperor โ€” issue edicts, your AI ministers execute.

multi-agentไธ‰็œๅ…ญ้ƒจdashboardaudit-trail12K-stars
Python
ClawX

ClawX

7.5K

Desktop app providing a graphical interface for OpenClaw AI agents. Turns CLI-based AI orchestration into a visual desktop experience without using the terminal.

desktopGUIagent-orchestrationvisual
TypeScript
OpenSpace

OpenSpace

6.8K

Agent optimization framework from HKUDS (NanoBot team). Makes agents smarter, lower-cost, and self-evolving. Useful for improving any science agent's performance over time.

HKUDSself-evolvingoptimizationlow-cost
Python
Open Multi-Agent

Open Multi-Agent

6.6K

TypeScript multi-agent framework. One runTeam() call from goal to result. Auto task decomposition, parallel execution. 3 dependencies, deploys anywhere Node.js runs. 4K stars in 4 days.

multi-agentTypeScriptlightweightparallel-execution
TypeScript
ClawTeam

ClawTeam

5.4K

Agent Swarm Intelligence โ€” one command, full automation. Agents spawn swarms, delegate tasks, deliver results. File + ZeroMQ P2P transport. Compatible with Claude Code, Codex, OpenClaw, nanobot, Cursor, any CLI agent. From HKUDS. 4.8K stars.

HKUDSmulti-agentswarmcoordination
Python
HiClaw

HiClaw

5.1K

Collaborative Multi-Agent OS by Alibaba. Transparent, human-in-the-loop task coordination via Matrix rooms. Open-source agent team orchestration for enterprise workflows.

Alibabamulti-agentMatrixhuman-in-the-loopenterprise
Shell
Clawith

Clawith

4.1K

OpenClaw for Teams. Multi-agent collaboration with persistent identity, long-term memory, and autonomous 'Aware' consciousness system. 6 trigger types: cron, once, interval, poll, on_message, webhook.

teamsmulti-agentawarenessFastAPI
Python
TinyAGI

TinyAGI

3.6K

Agent teams orchestrator for One Person Company. Coordinate multiple specialized agents for solo founders and indie researchers. Formerly TinyClaw.

orchestratorone-person-companymulti-agentindie
TypeScript
MetaClaw

MetaClaw

3.5K

Self-learning agent that evolves from every conversation. LoRA-based continual learning + reinforcement learning, no GPU cluster needed. Supports OpenClaw, NanoBot, PicoClaw backends.

meta-learningLoRAself-evolvingRL
Python
ClawManager

ClawManager

1.9K

Kubernetes-first control plane for managing OpenClaw and Linux desktop runtimes at team and cluster scale. Web UI for multi-instance orchestration.

Kubernetescluster-managementcontrol-planeweb-UI
TypeScript
ClawTeam-OpenClaw

ClawTeam-OpenClaw

1.4K

ClawTeam fork with deep OpenClaw integration. Default openclaw agent, per-agent session isolation, exec approval auto-config, production-hardened spawn backends. All upstream fixes synced.

multi-agentswarmOpenClaw-nativefork
Python
AlphaClaw

AlphaClaw

1.4K

The ultimate OpenClaw harness. Browser-based setup wizard, self-healing watchdog, Git-backed rollback, multi-agent management, prompt hardening, and one-click Railway/Render deploy. 892 stars, 440 tests. Zero SSH rescue missions.

devopsdeploymentwatchdogmulti-agentobservability
JavaScript
ArgusBot

ArgusBot

308

24/7 supervisor Agent for Codex CLI and Claude Code CLI. Keeps agents running, reviewing, and planning until the job is actually done. Useful for overnight experiments and long-running analysis tasks.

supervisor24/7overnightagent-management
Python
MagiClaw

MagiClaw

129

Conversational command center for scientific agent teams. From SJTU SAI Agents. Orchestrate multiple research agents through natural language dialogue.

SJTUcommand-centermulti-agentconversational
Python

Biomedicine & Omics

Medical & Clinical

EvoScientist

EvoScientist

4.3K

Self-evolving multi-agent AI research system. Harness Vibe Research โ€” 6 sub-agents (plan, research, code, debug, analyze, write) that evolve their capabilities over time. Built on DeepAgents framework. PyPI installable. Multi-platform: CLI, Telegram, Slack, WeChat. Apache 2.0 license.

AI4Sciencemulti-agentauto-researchaward-winning
Python
LabClaw

LabClaw

1K

Operating layer for LabOS โ€” Stanford-Princeton AI Co-Scientists project. Automated scientific co-discovery workflows for laboratory research.

StanfordPrincetonlab-automationco-discovery
Python
ClawBio

ClawBio

1K

The first bioinformatics-native AI agent skill library. Local-first, reproducible, built on OpenClaw. Covers genomics, population genetics, and equity-focused bioinformatics workflows.

bioinformaticsgenomicslocal-firstreproducible
Python
MedgeClaw

MedgeClaw

667

AI research assistant for biomedicine built on Claude Code with 140 K-Dense scientific skills. Real-time research dashboard, RStudio (:8787), JupyterLab (:8888), and Feishu rich card output.

biomedicineRNA-seqdrug-discoveryclinical
TeX
BioClaw

BioClaw

397

AI-powered bioinformatics research assistant built on OpenClaw. Integrates PubMed literature search and PyMOL molecular visualization for streamlined biological research workflows.

bioinformaticsPubMedPyMOLresearch-assistant
TypeScript
HealthClaw

HealthClaw

157

Open-source self-evolving personal health copilot. Medical consultation support, meal planning, report interpretation, chronic-care follow-up, wearable data analytics. Supports clinical data, medical imaging, and omics evidence. Streamlit + Feishu + CLI.

healthmedicalself-evolvingwearableomics
Python
OmicsClaw

OmicsClaw

154

Local-first cross-session AI assistant for life science researchers. 63 built-in skills across 6 omics domains (spatial transcriptomics, scRNA-seq, bulk RNA-seq, genomics, proteomics, metabolomics). Multi-Provider (Anthropic/OpenAI/DeepSeek/local LLM), Multi-Channel (CLI + Feishu/WeChat/Telegram sync), MCP + dynamic GitHub skill install. Persistent memory. 100% local โ€” no data leaves your machine.

6-omics63-skillsmulti-providerlocal-firstcross-session-memory
Python
STELLA

STELLA

146

Self-evolving multimodal agent platform for biomedical research. From Princeton/Stanford/UCLA. Published on bioRxiv and arXiv. Live demo at stella-agent.com.

PrincetonStanfordmultimodalself-evolvingbioRxiv
Python
MedClaw (zteyesreal)

MedClaw (zteyesreal)

66

Early-stage medical AI project exploring agent-based approaches for clinical decision support and medical data analysis.

medicalclinicalearly-stage
Python

Bioinformatics & Omics

Biomni

Biomni

3.5K

General-purpose biomedical AI agent from Stanford SNAP Lab. Broad coverage across biomedical tasks with a unified agent architecture.

Stanfordbiomedicalgeneral-purposeSNAP-Lab
Python
PantheonOS

PantheonOS

473

Evolvable, distributed agent framework for data science from the dynamo/Spateo team. Specialized in single-cell and spatial transcriptomics workflows.

dynamo-teamsingle-cellspatial-transcriptomicsdistributed
Python
CellVoyager

CellVoyager

250

AI CompBio agent that autonomously analyzes biological data and generates new insights. From Stanford Zou Group. Published in Nature Methods. Specializes in scRNA-seq with self-debugging and hypothesis generation.

StanfordNature MethodsscRNA-seqautonomous-analysis
Python
AutoBA

AutoBA

234

An AI agent for fully automated multi-omic analyses. Supports RNA-seq, scRNA-seq, spatial transcriptomics, WGS/WES, and ChIP-seq with automated code repair.

multi-omicsRNA-seqspatial-transcriptomicsself-debugging
Python
Stable
SpatialAgent

SpatialAgent

196

An AI agent for spatial biology by Genentech. Processes multimodal spatial genomics data for research insights.

spatial-biologygenomicsmultimodalGenentech
Python
CRISPR-GPT

CRISPR-GPT

174

LLM agent for CRISPR genome engineering. Assists with guide RNA design, experiment planning, and protocol optimization.

CRISPRgene-editingguide-RNA
Python
Stable

BioMedAgent

124

Biomedical AI agent published in Nature Biomedical Engineering. Dormant since Feb 2025 but with top-tier publication backing.

Nature-BMEbiomedicaldormant
Python
BioMaster

BioMaster

102

AI agent system for various bioinformatics tasks with specialized agent coordination. From the same group as POPGENAGENT.

bioinformaticsmulti-agentnucleome
Python
Stable

NeuroClaw

76

Neuroscience AI agent from CUHK AIM Group. Early stage โ€” no description yet, actively developed.

CUHKneuroscienceearly-stage
Python

Bioinfor-Claw

64

24/7 bioinformatics copilot โ€” 50 specific skills across 10 application scenarios (data access, multi-omics, CRISPR, gene-list, structure, ML, plotting, literature, lab tracking). Dual nature: standalone agent with browser-based chat UI + modular SKILL.md library that plugs into OpenClaw, Claude Code, or any custom agent. Multi-LLM backend (Anthropic, OpenAI, Google, Mistral, MiniMax, OpenAI-compatible).

bioinformaticscopilotskill-libraryopenclawmulti-LLM
HTML

Drug Discovery & Molecular

ChemCrow

ChemCrow

936

Pioneering chemistry AI agent (2023) from ur-whitelab. LangChain-based, cited in official docs. The foundational project that inspired MDCrow and the broader science agent movement.

chemistrypioneerLangChainur-whitelab
Python
Stable
Virtual Lab

Virtual Lab

712

A virtual lab of LLM agents for science research. Multi-agent system for nanobody design with principal investigator agent coordination.

multi-agentnanobody-designprotein-engineeringStanford
Python
Stable

BloClaw

337

Omniscient multi-modal agentic workspace for scientific discovery. Cheminformatics (RDKit), protein folding (ESMFold), molecular docking, autonomous RAG. XML-Regex dual-track routing, runtime sandbox for data visualization. From Beijing 1st Biotech + PLA General Hospital.

multi-modalcheminformaticsprotein-foldingmolecular-dockingRAG
Python
MDCrow

MDCrow

244

Molecular dynamics simulations with an LLM agent. Automates simulation setup, execution, and analysis using natural language.

molecular-dynamicssimulationcomputational-chemistry
Python
Stable
OriGene

OriGene

217

Self-evolving virtual disease biologist for mechanism-guided therapeutic target discovery. Automates disease mechanism analysis and drug target identification.

target-discoverydisease-mechanismself-evolvingdrug-target
Python
Stable
ChemGraph

ChemGraph

135

Agentic framework for computational chemistry and materials science workflows. From Argonne National Laboratory (ALCF). Orchestrates chemistry simulation pipelines.

Argonnechemistrymaterials-sciencesimulation
Python

DrugClaw (QSong)

132

Agentic RAG for drug intelligence: 57 skills, 15 task categories โ€” DTI, ADR, DDI, pharmacogenomics, repurposing. Powered by LangGraph.

57-skillsLangGraphDTIdrug-repurposing
Python

DrugClaw (DrugClaw)

127

AI Research Assistant for Accelerated Drug Discovery. Built with Rust for high performance. Covers compound screening, ADMET prediction, and lead optimization workflows.

Rustdrug-discoveryperformance
Rust
AIAgents4Pharma

AIAgents4Pharma

93

Multi-agent platform for pharmaceutical R&D. Covers drug discovery, drug development, and clinical pipeline automation. Built by VirtualPatientEngine.

pharmadrug-discoveryLangGraphknowledge-graph
Python

ChemClaw

51

Curated skills package for chemistry-focused AI workflows. Installable modularly via npx with stable and experimental skill tiers.

chemistryAI4Chem
Python

General Research

Autonomous End-to-End Research

autoresearch

91.7K

By Andrej Karpathy. Give an AI agent a small but real LLM training setup and let it experiment autonomously overnight. Single GPU, ~100 experiments/night. Works with Claude Code, Codex, or any agent.

KarpathyautonomousLLM-trainingsingle-GPU
Python

DeepResearch

19.7K

Tongyi Deep Research by Alibaba. Leading open-source deep research agent (30.5B params, 3.3B activated). State-of-the-art on Humanity's Last Exam, BrowseComp, and 4 more benchmarks. ReAct and IterResearch inference modes.

AlibabaTongyideep-researchbenchmark-SOTA
Python
Stable

AI-Scientist (v1)

14.3K

The original fully automated AI scientist from Sakana AI. Autonomously generates ideas, runs experiments, writes research papers, and conducts peer review. First open-source system to close the full research loop. 13K+ stars.

Sakana-AIend-to-endautonomouspaper-generation
Jupyter Notebook
Stable
RD-Agent

RD-Agent

14K

Microsoft R&D automation platform. AI-driven data and model research โ€” automates hypothesis generation, experiment design, and iterative improvement for industrial and scientific R&D.

MicrosoftR&D-automationdata-drivenindustrial
Python

pi-autoresearch

7.2K

Generic version of Karpathy's autoresearch loop that works on any measurable optimization target. Modify โ†’ run โ†’ score โ†’ keep/discard โ†’ repeat. Not limited to ML โ€” works for any codebase with a metric.

autoresearchoptimization-loopgenericany-metric
TypeScript
AI-Scientist-v2

AI-Scientist-v2

6.9K

Workshop-level automated scientific discovery via agentic tree search. By Sakana AI. Autonomously generates research papers through iterative hypothesis testing and experimentation.

Sakana-AIpaper-generationagentic-tree-searchautonomous
Python
Stable

AgentLaboratory

5.8K

End-to-end autonomous research workflow: Literature Review โ†’ Experimentation โ†’ Report Writing. Specialized LLM agents collaborate via arXiv, HuggingFace, Python, LaTeX. Includes AgentRxiv for cumulative cross-agent research.

multi-agentend-to-endAgentRxivautonomous
Python
Stable
AI-Researcher

AI-Researcher

5.6K

Autonomous scientific innovation system from HKUDS. NeurIPS 2025 paper. End-to-end research automation: literature review, hypothesis generation, experiment design, and paper writing. Production version at novix.science.

NeurIPS-2025autonomouspaper-writingHKUDS
Python
Stable

autoresearch (uditgoenka)

5.4K

Multi-purpose Claude Code plugin: autonomous goal-directed iteration inspired by Karpathy's autoresearch. Modify โ†’ Verify โ†’ Keep/Discard โ†’ Repeat. Also includes debug, fix, security audit, ship, and predict modes.

Claude-Codeautoresearchmulti-modeskill
Shell

OpenScience

๐Ÿ”ฅ New
2.7K

Open-source, model-agnostic AI workbench that runs the full research loop from literature review and hypothesis formation through code, experiments, analysis, and write-up. Includes specialist research agents, 290+ skills, scientific database tools, and a browser workspace.

end-to-end-researchmodel-agnosticresearch-agents290+-skillsscientific-computing
TypeScript
DATAGEN

DATAGEN

1.8K

AI-driven multi-agent research assistant automating hypothesis generation, data analysis, and report writing. Full research pipeline from question to publication-ready report.

hypothesis-generationdata-analysisreport-writingmulti-agent
Python

autoresearch-mlx

1.7K

Karpathy's autoresearch loop adapted for Apple Silicon Macs using MLX. No PyTorch required โ€” runs natively on M-series chips. Overnight autonomous experiment optimization.

Apple-SiliconMLXautoresearchMac-native
Python

Auto-Deep-Research

1.6K

Open-source, cost-efficient alternative to OpenAI's Deep Research, built on AutoAgent framework by HKUDS. Universal LLM support (OpenAI, Anthropic, DeepSeek, vLLM, Grok). One-click launch.

HKUDSdeep-researchcost-efficientuniversal-LLM
Python
Stable
NanoResearch

NanoResearch

1.5K

End-to-end autonomous AI research engine that actually runs experiments. 9-stage pipeline: idea โ†’ literature โ†’ code โ†’ GPU/SLURM training โ†’ real results โ†’ figures โ†’ LaTeX paper. Every data point, table, and chart comes from actual experiments, not LLM fabrication. CLI with TUI, Claude Code mode, Feishu bot. Local or cluster execution.

NanoBot-basedautonomousresearch-assistant
Python

InternAgent

1.4K

Unified long-horizon scientist framework from InternScience. Links deep research, executable verification, and memory-driven evolution across algorithm and empirical discovery. v1.5 with Web UI.

InternSciencelong-horizonend-to-endmemory-evolution
Python

OpenResearcher

1.1K

Fully open agentic LLM (30B-A3B) for long-horizon deep research by TIGER-AI-Lab. 54.8% on BrowseComp-Plus, surpassing GPT-4.1, Claude-Opus-4, Gemini-2.5-Pro. Open dataset (96K trajectories), model, and training code. Adopted by NVIDIA Nemotron.

TIGER-AI-Labopen-sourceBrowseCompNVIDIA-Nemotron
Python
Prismer

Prismer

783

Open-source alternative to OpenAI's Prism. Unified platform for AI-powered paper reading, learning, and research knowledge management.

open-sourcepaper-readingknowledge-managementPrism-alternative
TypeScript
EurekaClaw

EurekaClaw

699

Self-evolving AI research agent by the EurekaClaw team. Autonomous scientific discovery with persistent memory and tool-use capabilities.

self-evolvingautonomousdiscovery
Python

freephdlabor

578

Personalized multi-agent system that researches 24/7 on your scientific problem. From hypothesis generation through experimentation to publication-ready manuscripts. Dynamic workflows, fully customizable agents, human-in-the-loop.

24/7-researchmulti-agentcustomizablepaper-generation
Python
Stable

Kosmos

554

Autonomous discovery engine that tests hypotheses in sandboxed containers and tracks findings in a knowledge graph. Driven by Claude Code or API. Based on the Kosmos AI paper (arXiv:2511.02824).

autonomous-discoveryknowledge-graphsandboxedhypothesis-testing
Python

InnoClaw

384

A self-hostable AI research workspace with grounded chat, paper study tools, and scientific skill execution โ€” all anchored to your local files.

innovationSpectrAI
TypeScript

OmniScientist

155

AI Scientist Ecosystem by Tsinghua FIB Lab and Zhongguancun Academy. 7 core papers covering co-evolving ecosystem, expert knowledge integration, high-impact research question identification, and automated experiment design.

Tsinghuaecosystemmulti-paperfull-process
Python
Stable

NeuriCo

147

Autonomous research framework from University of Chicago Human+AI Lab. Input a title, domain, and hypothesis โ€” agents handle literature review, experiment design, code execution, analysis, and LaTeX paper writing. Supports Claude Code, Codex, and Gemini CLI. IdeaHub platform for community research ideas. Docker-ready, domain-agnostic.

โšœ๏ธ Editor's Pick โšœ๏ธ

Beautifully minimal design: just a hypothesis in, full paper out. The IdeaHub integration is clever โ€” a community platform where anyone can submit research ideas for agents to explore. One of the few frameworks that truly treats the AI as a co-scientist, not just a code generator.

autonomous-researchmulti-providerIdeaHubUChicagopaper-writing
Python
Open Co-Scientist

Open Co-Scientist

51

Open-source adaptation of Google DeepMind's AI Co-Scientist. Generates, reviews, ranks, and evolves research hypotheses using multi-agent architecture.

Google-Research-inspiredhypothesis-evolutionmulti-agentopen-source
Python
Stable

Personal Research Workbench

ARIS

13.7K

ARIS v0.4.1 โ€” Lightweight autonomous ML research skills. 62 skills for cross-model review loops, idea discovery, experiment automation. Now with standalone CLI, Plan mode, Research Wiki (persistent knowledge base), Self-Evolution (/meta-optimize), local model support (LM Studio/Ollama). Works with Claude Code, Codex, OpenClaw, any LLM agent. No framework lock-in.

Claude-Codesleep-researchcross-modelMarkdown-skills
Python
DeepScientist

DeepScientist

3.2K

Local-first AI research studio. 15-minute setup, one repo per research quest, visible progress, human takeover anytime. Published at ICLR 2026.

ICLR-2026local-firstresearch-studiohuman-in-the-loop
TypeScript
Memento-Skills

Memento-Skills

1.5K

Let Agents Design Agents. A meta-framework where AI agents create, test, and evolve new skills and agents autonomously. Unlike traditional skill libraries, the skills here are machine-authored.

โšœ๏ธ Editor's Pick โšœ๏ธ

A genuinely novel approach: instead of humans writing skills for agents, agents write skills for agents. Worth watching as a preview of where the entire ecosystem is heading.

meta-frameworkagent-designs-agentself-evolvingautonomous
Python
Dr. Claw

Dr. Claw

1K

AI Research IDE with built-in "AI Doctors" assistants. Full research lifecycle: Survey โ†’ Ideation โ†’ Experiment โ†’ Publication โ†’ Promotion. 100+ research skills, auto pipeline planning, real-time task tracking, LaTeX rendering, Git integration. PWA-ready.

research-IDEmulti-agentLaTeXPWAfull-lifecycle
JavaScript
ScienceClaw (beita6969)

ScienceClaw (beita6969)

868

Self-evolving AI research colleague. 285 skills that grow at runtime โ€” the agent writes new skills based on your usage patterns. 4-layer persistent memory with LanceDB, 1hr+ session timeout, zero-hallucination citation protocol.

self-evolvingpersistent-memoryzero-hallucination28-disciplines
TypeScript
ScienceClaw (TaichuAI)

ScienceClaw (TaichuAI)

561

Personal research assistant built with LangChain DeepAgents and AIO Sandbox. A new architecture beyond OpenClaw, offering stronger security, better transparency, and a more user-friendly experience.

LangChainDeepAgentsbeyond-OpenClawsecuritytransparency
Python

CodeScientist

345

Automated scientific discovery system by Allen AI. LLM-as-a-mutator generates experiment ideas from paper+code combinations, then creates, runs, and debugs experiments in containers. Writes reports and meta-analyses.

Allen-AIcode-experimentsgenetic-mutationdiscovery
Python
OpenLens AI

OpenLens AI

272

Fully autonomous multimodal research agent. Combines vision and text understanding for scientific discovery. Chinese academic origin with bilingual support.

multimodalautonomousvision+textbilingual
Python

Wisp Science

๐Ÿ”ฅ New
258

Open-source, local-first desktop AI research workbench for scientific computing. Runs persistent Python and R environments locally or on SSH, WSL, and GPU compute; loads Agent Skills and connects to around 80 bioinformatics databases through bundled MCP servers.

local-firstdesktop-workbenchPython-Rbioinformatics-MCPreproducible-research
Rust
ScienceClaw (MIT LAMM)

ScienceClaw (MIT LAMM)

235

Open-source AI agent swarm for crowdsourced scientific discovery. 300+ interoperable tools, plannerless ArtifactReactor coordination, DAG lineage tracking. Agents self-coordinate across institutions via Infinite platform (lamm.mit.edu/infinite). Solving real problems: peptide binder design, ceramic discovery, cross-disciplinary pattern recognition. From MIT LAMM Lab (Markus Buehler). Genesis Mission backed.

MITagent-swarmInfinitemulti-agentcrowdsourced-discovery
Python
BioAgents

BioAgents

178

AI scientist framework for autonomous deep research in biological sciences. Multi-agent system combining literature analysis with data scientist agents for iterative scientific discovery.

biologyliterature-analysisdata-scienceiterative-discovery
TypeScript

DrClaw (InternScience)

169

A 7x24 AI research team for academics with 244 skills and 13 agent templates, designed to eliminate repetitive lab work so researchers can focus on thinking.

InternScience
Python
BioDiscoveryAgent

BioDiscoveryAgent

114

LLM-based AI agent for closed-loop design of genetic perturbation experiments. Autonomous reasoning for biological discovery.

genetic-perturbationexperiment-designStanford
Python
Stable
StatsClaw

StatsClaw

87

Multi-agent workflow for statistical software development. Information-isolated agents: builder (no ground truth), simulator (no algorithm), tester (independent validation). Cambridge + Stanford.

CambridgeStanfordstatisticsR-packagesmulti-agent
Python
SciClaw (drpedapati)

SciClaw (drpedapati)

85

Paired-scientist AI assistant for reproducible research. Lightweight Go runtime, lifecycle hooks, manuscript integration, 12 baseline scientific skills. PicoClaw-compatible. Homebrew installable.

GoreproduciblePicoClaw-compatibleHomebrewpaired-scientist
Go

MatClaw

80

Open materials-science agent that turns natural-language tasks into reproducible simulation workflows.

materials-sciencesimulation
TypeScript
ScienceClaw (Zaoqu-Liu)

ScienceClaw (Zaoqu-Liu)

58

AI Research Lab That Never Sleeps. 9 agents, 263 skills, 77 databases โ€” from literature to publication, every discipline, zero boundaries. Full research lifecycle automation.

9-agents263-skills77-databasesfull-lifecyclemulti-discipline
Shell
SR-Scientist

SR-Scientist

53

Scientific equation discovery with agentic AI. Published at ICLR 2026. Symbolic regression via AI agents from GAIR Lab (Shanghai Jiao Tong University).

ICLR-2026equation-discoverysymbolic-regressionSJTU
Python
Stable

Claude Science

๐Ÿ”ฅ New

Anthropic's integrated AI workbench for scientists. A coordinating agent uses 60+ curated skills and connectors across genomics, single-cell, proteomics, structural biology, and cheminformatics, with local, SSH, and HPC compute plus auditable figures and manuscripts. Closed-source beta for Claude Pro, Max, Team, and Enterprise.

Anthropicscientific-workbench60+-skillsHPCreproducible-artifacts

Paper & Literature Tools

AutoResearchClaw

AutoResearchClaw

13.8K

Fully autonomous, self-evolving research framework: from a one-line idea to a conference-level paper. 23-stage pipeline with multi-agent debate, self-healing error correction, citation verification, and hardware-adaptive execution. Compatible with OpenClaw and MetaClaw.

idea-to-paperself-evolving23-stagecitation-verificationOpenClaw-compatible
Python

Research-Claw

845

Full-stack AI research assistant โ€” "In the AI era, why can't everyone be a PI?" Supports extensible skills, literature review, note-taking, experiment tracking, and paper writing end to end. 350+ plugins available.

full-stackPI-for-everyoneextensible
TypeScript

ResearchClaw (ymx10086)

310

Personal AI assistant built for research. Fast setup, local or cloud, integrates with chat apps. Streamlines literature review, note-taking, experiment tracking, and paper writing end to end.

end-to-endlocal+cloudchat-integration
Python

CitationClaw

310

Turning every citation into explainable impact. Citation analysis and impact assessment tool for academic research โ€” tracks citation networks, measures influence, and provides explainable bibliometric insights.

citation-analysisimpactexplainable
Python

research-claw (nanoAgentTeam)

285

Self-hosted AI assistant for academic research โ€” manages papers, searches literature, tracks deadlines.

self-hosteddeadlines
Python

PaperClaw (guhaohao0991)

242

OpenClaw skill for paper search-review-critique expert-agent. Generates research workflows for specific topics (demo: Scientific ML and 3D geometry).

paper-reviewexpert-agentSciML
Python

RS-PaperClaw

164

Daily-automated remote sensing paper pipeline. Pulls arXiv candidates, applies keyword + LLM filtering, generates per-paper GitHub Issues as reading reports, compiles daily digest Issues, archives to markdown, and pushes to DingTalk/Feishu. Live reader on GitHub Pages.

remote-sensingpaper-digestautomationarxivgithub-issues
Python

ClawPhD

155

Agent that turns academic papers into publication-ready diagrams, posters, videos, and more. Visual research communication tool โ€” automates the creation of conference materials from your manuscripts.

diagramspostersvideosvisualization
Python

PaperClaw (meowscles69) โ˜ ๏ธ

141

[REMOVED] GitHub account deleted. Was: 27 OpenClaw skills for academic research teams โ€” literature reviews, hypothesis versioning, grant writing, lab knowledge handoffs. 141 stars at time of removal.

โšœ๏ธ Editor's Pick โšœ๏ธ

GitHub account meowscles69 was deleted on or before 2026-04-13. Repo and all code lost.

removedaccount-deleted
Python
Dormant

PaperClaw (1692775560)

70

End-to-end 23-stage AI pipeline for academic research. Three-layer architecture โ€” LLM decision brain, skill execution layer, memory knowledge management. Ships with a modern dark-theme Web UI (one-click research start, real-time stage tracking, in-browser result preview) plus full CLI. 1,634 tests passing.

paper-analysis23-stage-pipelineweb-uithree-layer-architecture
Python

ResearchClaw (Noietch)

65

Standalone Electron desktop app combining AI-powered paper management, interactive reading notes, and research idea generation โ€” no browser or server required.

desktop-appTypeScript
TypeScript

PaperClaw (hurtjan)

3

Transform your PDF paper library into a cross-referenced, queryable citation graph database. AI agents extract metadata, citations, and context from PDFs, then fuzzy-match across papers. Dual environment: Main (ingest) + Query (Haiku-powered, read-only). 41 commits, actively developed.

citation-graphPDF-ingestionliterature-databaseClaude-Code
Python
Early Stage

Education

OpenMAIC

OpenMAIC

19.9K

Open Multi-Agent Interactive Classroom by Tsinghua University. One-click immersive AI classroom with multi-agent learning experience. Published in JCST'26. Live demo at open.maic.chat. Built with Next.js 16 + React 19 + LangGraph. Vercel one-click deploy. OpenClaw integration. 15K+ stars.

Tsinghuaclassroommulti-agentPBL
TypeScript

MathClaw

443

Multimodal math learning assistant for middle and high school. Solving workspace, weakness diagnosis, knowledge graphs, error graphs, auto-summaries. Supports WeChat, QQ, Feishu, Telegram, Discord. From Renmin University.

educationmathematicsmultimodalknowledge-graphRenmin-University
Python

Benchmarks & Evaluation

PinchBench

1.3K

Real-world benchmarks for AI coding agents. 23 tasks across productivity, research, writing, coding, analysis, email, memory, and skills. Public leaderboard at pinchbench.com. Auto-grading + LLM judge. Tests what actually matters: tool usage, multi-step reasoning, and practical outcomes.

benchmarkevaluationcoding-agentleaderboardOpenClaw
Python

Claw-Eval

728

End-to-end evaluation suite for autonomous agents. 300 human-verified tasks across 9 categories (service orchestration, multimodal perception, multi-turn dialogue). Trajectory-aware grading over 2,159 rubrics. Tests Completion, Safety, and Robustness. Used by Qwen, GLM, MiniMax.

benchmarkevaluationmultimodaltrajectory-awaresafety
Python

ClawBench

510

Real-world web task benchmark for AI agents. 153 tasks across 15 life categories on 144 live websites. Claude Sonnet 4.6 achieves only 33.3%, GPT-5.4 only 6.5% โ€” highlighting the gap between sandbox and real-world performance. From UBC, Vector Institute, CMU, SJTU, Tsinghua.

benchmarkweb-tasksreal-world153-tasksmulti-university
Python

ResearchClawBench

221

Benchmark for evaluating whether AI agents can independently conduct scientific research. 40 real-science tasks across 10 disciplines (Astronomy, Chemistry, Physics, Life, etc.), two-stage pipeline (autonomous research + LLM peer-review scoring), expert-annotated checklists. Score 50 = match human paper, 70+ = surpass it. Supports Claude Code, Codex CLI, OpenClaw, NanoBot, EvoScientist. From Shanghai Jiao Tong University.

benchmarkevaluationmulti-agentscientific-researchSJTU
Jupyter Notebook

BixBench

132

Comprehensive benchmark for LLM-based agents in computational biology. 205 questions extracted from 60 real-world published Jupyter notebooks (capsules). Tests dataset exploration, multi-step analysis (Python/R/Bash code execution), and scientific hypothesis generation. Supports both agentic and zero-shot evaluation modes. From Future House (PaperQA / Aviary lab).

benchmarkcomputational-biologyjupyter-notebookreal-world-dataFuture-Houseagentic-eval
Python
Stable

ClawMark

117

Multi-day, multimodal coworker agent benchmark. 100 tasks across 13 professional domains (insurance, legal, EDA, etc). Environment changes dynamically โ€” new emails, file updates, schedule shifts. Cross-modal multi-turn with branching logic. Best score only 55%. From NUS, Evolvent AI, HKU, MIT, UW, UC Berkeley, CUHK, HKUST (40+ scholars).

benchmarkmulti-daymultimodaldynamic-environmentNUS-MIT-Berkeley
Python

BioAgent Bench

31

AI agent evaluation suite for bioinformatics. Benchmarks LLM agents on real bioinformatics tasks. Covers sequence analysis, genomics workflows, and computational biology pipelines. arXiv paper.

benchmarkbioinformaticsevaluationLLM-agent
Python
Stable

ClawSafety (Benchmark)

18

Safety benchmark for personal AI agents under realistic prompt injection. 120 adversarial test cases across 5 harm domains, 3 attack vectors, and 5 harmful action types. Tested Claude, Gemini, GPT-5.1, DeepSeek on OpenClaw/Nanobot/NemoClaw scaffolds. Key finding: chat safety โ‰  agent safety.

safetybenchmarkprompt-injectionadversarial
Python
Early Stage

Security

ClawVault

ClawVault

1.2K

OpenClaw security vault for AI agents. Sensitive data detection (API keys, PII, credit cards, 15+ pattern types), prompt injection defense, dangerous command guard (rm -rf, curl|bash), auto-sanitisation with placeholder restoration, token budget control, real-time web dashboard. Transparent proxy gateway sits between AI tools and external APIs. From Tophant.

securityprompt-injectionsensitive-dataauditvault
Python

ClawSafety (Scanner)

2

Security scanner for Agent Skills โ€” the npm audit for the Agent-Native ecosystem. Scans for shell injection, hardcoded secrets, unpinned dependencies, excessive permissions, and prompt injection risks. Rust CLI with GitHub App integration. Inspired by the ClawHavoc incident (341 malicious skills).

securityscannerskillsvulnerability
Rust
Early Stage

Skill Evolution Frameworks

SkillOpt

SkillOpt

13.8K

Text-space optimizer from Microsoft Research that trains reusable natural-language skills for frozen LLM agents โ€” like training a neural net but on a Markdown document. Trajectory-driven add/delete/replace edits, validation-gated updates, textual learning-rate budget. Produces a deployable best_skill.md (300-2,000 tokens). Best or tied-best on all 52 (model ร— benchmark ร— harness) cells; lifts GPT-5.5 +19.1 inside Claude Code.

Microsoftskill-evolutiontext-space-optimizervalidation-gated
Python
TextGrad

TextGrad

3.7K

Automatic differentiation via text โ€” uses LLMs to backpropagate textual gradients across prompts, skills, queries, and code. From Stanford Zou Group (same lab as Virtual Lab, CellVoyager). Published in Nature. The foundational primitive that newer skill-evolution frameworks (SkillOpt, SkillClaw, DSPy optimizers) compose on top of.

StanfordNature-papertext-autogradprimitive
Python
Stable
SkillClaw

SkillClaw

2.2K

Collective skill evolution framework from Alibaba DreamX Team (AMAP-ML). Aggregates real-user trajectories across sessions, agents, devices, and users, then runs an autonomous evolver to identify recurring patterns and refine the shared skill library โ€” improvements discovered in one context propagate to everyone. Compatible with Hermes, OpenClaw, Codex, Claude Code, QwenPaw, IronClaw, PicoClaw, ZeroClaw, NanoClaw, NemoClaw, and any OpenAI-compatible API. Validated on WildClawBench; significantly improves Qwen3-Max.

AlibabaDreamXcollective-evolutionmulti-user-traces
Python

๐Ÿš€

Missing the project in your mind?

Know an AI agent project for scientific research that should be listed here? Submit it and we'll list it!

Submit Your Project

Takes 30 seconds ยท Just a GitHub link + one-line description

The Claw Ecosystem in Numbers

A rapidly growing family of AI agent projects powering scientific research worldwide.

152+ Projects Tracked

152+

Projects Tracked

2100+ Skills Indexed

2100+

Skills Indexed

12 Project Categories

12

Project Categories

Frequently Asked Questions

Common questions about the OpenClaw science ecosystem.










Claw4Science - AI Agent Directory for Scientific Research