AI coding has made CI a bottleneck, so we reworked ours to keep up
Thread
Unofficial Hacker News client; not affiliated with Y Combinator.
AI coding has made CI a bottleneck, so we reworked ours to keep up
Unofficial Hacker News client; not affiliated with Y Combinator.
torben-friis · · focus · HN ↗
Everyone's going so fast that they keep hitting walls. Review, CI, product asking for things, whatever.
Why have we not seen an improvements in products?
While every post and thread feels like a 90's wall street office, the new android and iphone ship with fewer features than usual. No indie guys come up with a linux-sized alternative OS. Switch 2 remains unhacked. Windows takes 3 seconds to show the right click menu.
Is everyone just running full speed in circles or something?
saithound · · focus · HN ↗
The simplest explanation is that they don't give a flying flamingo about what you or I consider "improvements to products".
This report is an example.
There are several changes that modify CI behaviour, where the article gives no corresponding quality measurement.
They replaced type aware custom lint rules with AST-only static analysis. They don't say anything about what those new rules detect, didn't do old-vs-new rule comparison. They switched the TypeScript check from tsc to tsgo. Again, they are very proud of the performance improvement, but don't seem to care about diagnostic equivalence. The list goes on. They don't even report pass/fail agreement between the old and the new CI. They have 4x more tests, but no idea whether this big test suite works any better than the smaller old one, or even whether it works at all.
plant-ian · · focus · HN ↗
Aeolun · · focus · HN ↗
Rapzid · · focus · HN ↗
And dumping tests it needs for intermediate work steps in your suite to run for all of eternity.. Sometimes it'll create these in tmp, but not always!
plagiarist · · focus · HN ↗
maccard · · focus · HN ↗
nottorp · · focus · HN ↗
Rapzid · · focus · HN ↗
Unless you prompt them otherwise, the models tend to write WAY too many useless tests and in the most inefficient ways imaginable. This balloons test counts and lines of code to insane levels and frankly, likely, slowly makes it more and more costly for the AI to make future changes.. To the point it can't wrap its context around the code base and effectively make necessary changes.