Flash News

React-specific measurements: GPT-5.6 Sol first, but best achieved only 43.1%

On July 16th, the Million team, which specializes in the React development tool, launched ReactBench v1, to test whether AI programming Agent can perform a real React development and avoid introducing new code questions. React is the main JavaScript library used to build websites and App interfaces. Million has previously developed open-source tools such as React Scan, React Doctor and Million.js. ReactBench sifts 51 real jobs from open-source projects. In addition to checking the functioning of the function, it uses over 400 rules to screen procedural errors, performance, accessibility and code quality issues. GPT-5.6 Sol received a combined separation of 43.1% and Fable 5 41.2%. Officially, the difference between the two is smaller and it is not possible to confirm that Sol is significantly stronger for the time being. The single test cost of Fable 5 is about 6.3 times that of Sol when it is also an XHigh configuration. Even with the highest-performing configuration, the task successfully accomplished is less than half. Of the 4455 new functional development tests, 1194 React problems were introduced in the models, of which 77.5 per cent were procedural errors or security problems。

OKX - Unlock Rewards