Skip to content
Navigation menu
Search
Powered by Algolia
Search
Log in
Create account
DEV Community
Close
← All Trends
Custom Local Evals for New AI Coding Models
77 posts in this trend in the last 7 days
•
Active about 3 hours ago
A New Model Dropped. Is It Actually Good at Your SQL? A 30-Minute Smoke Test
Morgan Li
Morgan Li
Morgan Li
Follow
Aug 10
A New Model Dropped. Is It Actually Good at Your SQL? A 30-Minute Smoke Test
#
ai
#
sql
#
testing
#
opensource
Comments
Add Comment
4 min read
Treat Every “Cheap and Great” Model Release as a Hypothesis: A Reproducible LLM Cost-Quality Router
Morgan Xu
Morgan Xu
Morgan Xu
Follow
Aug 13
Treat Every “Cheap and Great” Model Release as a Hypothesis: A Reproducible LLM Cost-Quality Router
#
ai
#
llm
#
python
#
devops
Comments
Add Comment
4 min read
The Week After the Eval: A Cost-Aware Routing Harness for New Model Drops
Riley Wu
Riley Wu
Riley Wu
Follow
Aug 13
The Week After the Eval: A Cost-Aware Routing Harness for New Model Drops
#
ai
#
llm
#
productivity
#
tooling
Comments
Add Comment
5 min read
Stop Asking Coding Models to Write Code. Test Whether They Can Review a Patch
Finley Zhou
Finley Zhou
Finley Zhou
Follow
Aug 13
Stop Asking Coding Models to Write Code. Test Whether They Can Review a Patch
#
ai
#
python
#
testing
#
codequality
Comments
Add Comment
5 min read
Stop Arguing About Which Model Is Best. Build a Two-Tier Habit Instead.
Avery Wang
Avery Wang
Avery Wang
Follow
Aug 13
Stop Arguing About Which Model Is Best. Build a Two-Tier Habit Instead.
#
ai
#
productivity
#
programming
#
tooling
Comments
Add Comment
5 min read
« First
‹ Prev
1
2
3
4
5
We're a place where coders share, stay up-to-date and grow their careers.
Log in
Create account