behat
- Karma
- 134
- Created
- 3 years ago
About
building tuneloop.ioRecent Submissions
- 1. ▲ A look at coding agent benchmarks, and what may be interesting next (tuneloop.io)
- 2. ▲ What would it take to match model intelligence to the task? (tuneloop.io)
- 3. ▲ Benchmarking – Frontier models go out of their way to cheat (tuneloop.io)
- 4. ▲ Show HN: Tuneloop – a local CLI for analyzing coding agent session transcripts (github.com)
- 5. ▲ Launch HN: Relvy (YC F24) – On-call runbooks, automated (relvy.ai)
- 6. ▲ Ramp: How we made Ramp sheets self-maintaining (twitter.com)
- 7. ▲ LLM Costs of AI investigating production alerts (relvy.ai)
- 8. ▲ OpenRCA benchmark – Improving Claude's root cause analysis accuracy by 12 pp (relvy.ai)
- 9. ▲ Can AI debug problem scenarios in the OpenTelemetry demo application? (relvy.ai)
- 10. ▲ How GitHub Copilot is getting better at understanding your code (github.blog)