Skip to content

Entry

SWE-agent

Appears in 6 awesome lists

SWE-agent takes a GitHub issue and tries to automatically fix it, using GPT-4, or your LM of choice. It solves 12.29% of bugs in the SWE-bench evaluation set and takes just 1.5 minutes to run.

Open github.comprinceton-nlp/swe-agent

Found in these lists

AI Game DevTools (AI-GDT)

Section: Game (World Model & Agent) · Agent Computer Interfaces Enable Software Engineering Language Models.

ActiveScore 77

Open-source projects

Section: Links

FreshScore 87

Awesome Ai Agents 2026

Section: Autonomous Software Engineers · Princeton. Resolves real GitHub issues autonomously.

ActiveScore 74

Awesome AI Coding Tools

Section: Coding Agents · Princeton's autonomous agent that resolves real GitHub issues by navigating repos, editing files, and running tests.

FreshScore 85

awesome-ChatGPT-repositories

Section: Others · SWE-agent takes a GitHub issue and tries to automatically fix it, using GPT-4, or your LM of choice. It solves 12.29% of bugs in the SWE-bench evaluation set and takes just 1.5 minutes to run.

FreshScore 87

Awesome AI Agents: Tools, Resources, and Projects

Section: Repositories · SWE-agent takes a GitHub issue and tries to automatically fix it, using GPT-4, or your LM of choice. github

SlowScore 68

LangChain

Langchain integrates various providers like Anthropic, AWS, and OpenAI, and offers tools for components such as LLMs, chat models, and data analysis, supporting functionalities from Alpha Vantage to YouTube github | docs

In 20 listsDetails

LlamaIndex

(MIT) provides modules for structured outputs at different levels of abstraction, including output parsers for text completion endpoints, Pydantic programs for mapping prompts to structured outputs using function calling or output parsing, and pre-defined Pydantic programs for specific output types.

In 14 listsDetails

Dify

February 2026 release making human oversight a native workflow primitive: suspend execution at critical decision points, expose review-and-edit UI mid-flow, and route subsequent execution based on human action (approve/reject/escalate). Demonstrates how HITL transitions from bolt-on approval gates…

In 14 listsDetails

AutoGen

Microsoft's multi-agent conversation framework with a complete AgentChat layer covering agent loop, tool integration, termination conditions, and human-in-the-loop. The most comprehensive open-source reference for large-scale multi-agent harness design.

In 14 listsDetails

Pipecat

Handles frame management, streaming media coordination, and pipeline orchestration between ASR/LLM/TTS services for sub-800ms Total Turn-Around Time voice interactions. The missing harness primitive for voice agents: manages backpressure, handles frame queueing, and exposes a simple async…

In 7 listsDetails

OpenHands

🙌 OpenHands: Code Less, Make More. (formerly OpenDevin), a platform for software development agents powered by AI. github

In 7 listsDetails

AgentBench

Multi-environment agent benchmark (OS, DB, web, code) with a structured eval pipeline. Worth studying for its environment isolation design and task definition format when building custom eval environments for your harness.

In 7 listsDetails

OpenAgents

The web-browsing agent module of the OpenAgents platform (HKU). Enables autonomous navigation of websites via natural language, as part of a larger multi-modal agent framework.

In 7 listsDetails