Benchmark Results
Topic archive • 2 matches
AI-generated: summaries written by AI from the linked sources. How we use AI
2026-10-10
Technology
Lenovo's TianxiCode coding agent, powered by DeepSeek-V4.1-Flash, achieved a 71% verified score on SWE-bench-Live Lite. This performance marks a significant milestone for in-house enterprise AI coding frameworks.
Benchmark Results • Pandaily
Permalink
2026-09-28
Technology
This video delivers the latest AI model news, including upcoming releases like Sonnet 5.5 and GPT-7 rumors, plus benchmark results for DeepSeek V5 and Qwen 4. It's a concise roundup for professionals who need to stay updated on the rapidly evolving AI landscape.
AI News • Universe of AI
Permalink