IBM Research has addressed the unreliability issue in artificial intelligence agents that pass tests on initial attempts but fail upon repetition. By evaluating decision uncertainty step by step, their new method identifies potential failure points and generates targeted guidelines to ensure repeatable execution.
Google has introduced Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking to advance real time voice dialogue. These additions bring parallel background reasoning, mid conversation multilingual switching, and instant tool calling to conversational AI agents.