AI Agents for a Researcher — Page 5
Controllable Complex RAG
Plans, retrieves, and verifies answers to complex multi-step questions over your own documents.
ASO & App Marketing Skills
AI agent skills for App Store Optimization and app marketing, providing expert-level guidance for Cursor, Claude Code, and any Agent Skills-compatible assistant.
Concordia: A Library for Generative Social Simulation
A library for constructing and running generative agent-based models that simulate social interactions.
BettaFish Public Opinion Research
Research public opinion across social media, the web, and private data, then produce interactive reports.
Meta ARE
Evaluate AI agents on evolving, real-world tasks that demand multi-step reasoning and adaptation.
Timeplus Proton: Streaming SQL Engine
A unified streaming SQL engine in a single C++ binary, delivering millisecond latency and 100+ GB/s throughput as a simpler, faster alternative to ksqlDB and Flink.
Karpathy Agentic ML Engineer
Automated ML engineer powered by Claude Agent SDK and Google ADK, leveraging Scientific Agent Skills to train state-of-the-art models.
ShoppingAgent Order Analyst
Captures order pages, extracts their contents with a configured model, and exports the results to Excel.
OpenBB Open Data Platform
The open-source data infrastructure for analysts, quants, and AI agents — connect once, consume everywhere.
GenoTEX Gene Expression Benchmark
An expert-curated benchmark for evaluating automated gene-expression and gene–trait association analysis workflows.
JARVIS / HuggingGPT
A ChatGPT-controlled system that plans tasks and orchestrates Hugging Face expert models for multi-step AI requests.
M-flow Cognitive Memory Engine
A graph-driven RAG paradigm that scores evidence paths for human-like associative recall.
NOF0 AI Trading Arena
Compare LLM and prompt trading strategies through live-market PNL.
DeepAnalyze
An autonomous data-science assistant for analyzing diverse data sources and producing professional reports.
RLinf RL Infrastructure
Scalable reinforcement-learning infrastructure for embodied robotics and agentic AI training.
Windows Agent Arena
Benchmark multimodal desktop agents in a reproducible Windows 11 environment, locally or at Azure ML scale.
QuantMind: Quantitative Finance Knowledge Engine
Transform raw financial information into structured, traceable, queryable knowledge.
ART Agent Reinforcement Trainer
Train multi-step LLM agents from rollout rewards with GRPO and LoRA updates.
Xiaohei Hand-Drawn Illustrations Skill
Turn judgments, flows, states, and metaphors in Chinese articles into 16:9 white-background hand-drawn quirky illustrations.
OpenLens Research Agent
Turn a dataset and one research idea into an autonomous multimodal research workflow.
OSWorld
Benchmark multimodal computer-use agents on open-ended tasks in real desktop environments.
PageIndex: Vectorless, Reasoning-based RAG
Build hierarchical tree indexes and reason over them for context-aware retrieval - no vector DB, no chunking.
ToolBench & ToolLLaMA
Train, run, and evaluate API-using models with real REST API data, retrieval, and DFSDT.
veRL Agent Training
Multi-turn reinforcement learning training for long-horizon LLM and vision-language agents.