Evaluate agent performance

https://storage.googleapis.com/gweb-cloudblog-publish/images/09_-_Data_Analytics_tFH57V6.max-2600x2600.jpg

Editor’s note: Some of the most interesting questions in AI are being asked by information theoreticians, around how to provide context to an emerging class of AI agents. A few weeks ago, we waded into those waters with a blog about the Open Knowledge Format, a specification that formalizes the LLM-wiki pattern into a portable, interoperable format to represent the metadata, context, and curated knowledge that modern AI systems need to operate. That blog generated a ton of interest, so we’ve decided to bring you more of the same, as part of our new “Frontier and Center” series. Today, we hear from two members of Google Data Cloud’s frontier AI team on the recurring challenge of how to systematically evaluate whether or not an agent is able to answer questions effectively based on its context. Read on for more, and watch this space for more blogs from this team.


...

Copyright of this story solely belongs to google.com. To see the full text click HERE

Read more

https://cdn.mos.cms.futurecdn.net/nbGANf34NLVwWVjzLTLQWg-1403-80.jpg

Major sodium-ion battery milestone passed by California startup — Unigrid aces Hyundai’s safety tests down to -20C as experts call tech a 'compelling alternative' to lithium-ion

Lithium-ion batteries are one of the most successful technologies ever invented, and now span everything from smartphones to power stations — but the technology is not without its flaws and limitations, and scientists are now busy trying to figure out which type of next-gen battery will take over next. One of