Apodex launches TRACES benchmark to evaluate AI for scientific discovery
Singapore, Aug. 18 -- Apodex has launched TRACES, a benchmark designed to evaluate how artificial intelligence systems perform on open-ended scientific problems where the correct answer may not yet be known.
Unlike conventional AI benchmarks that rely on static datasets and predefined answer keys, TRACES places AI systems inside executable environments where they can observe, act, use tools, receive feedback and revise their approach while working towards a verifiable outcome.
The benchmark is intended to assess what Apodex describes as "discoverative AI" - systems designed to identify new findings from existing knowledge rather than reproduce information already contained within training data.
Depending on the scientific problem, a TR...
Click here to read full article from source
इस लेख के रीप्रिंट को खरीदने या इस प्रकाशन का पूरा फ़ीड प्राप्त करने के लिए, कृपया
हमे संपर्क करें.