S
SWE-Touch
Benchmark for testing coding agents' handling of concurrent user edits in shared workspaces
Open SourceFree
About
Research benchmark framework that evaluates how coding agents respond when users modify code during active tasks. Introduces the Counter-Edits methodology to stress-test agent robustness in collaborative editing scenarios. Essential for developers building AI coding assistants that need to handle real-time workspace changes and maintain context during concurrent modifications.
Details
| Type | |
| Integrations | |
| Language |
Tags
coding-agentevaluationopen-sourceautonomous
Quick Info
- Organization
- Research Team
- Pricing
- open-source
- Free Tier
- Yes
- Updated
- Aug 4, 2026
Also in Dev Tools
C
Crawl4AI
Open-source web crawler optimized for LLMs and AI agents — 62K+ stars
OSSFree
unclecode
76.0K5d ago82
F
Firecrawl
Web scraping API built for LLMs — turn any website into LLM-ready data — 89K+ stars
OSSfreemium
Mendable
160.4Ktoday159
H
Headroom Context Optimization
Reduce LLM API costs by 50-90% through advanced context compression
OSSFree
Shubham Saboo
130.3Ktoday91