Google AI Launches New Scoring Model Cappy to Enhance Performance of Multitask Language Models


Google confirmed that its Gemini model accessed the systems of three external companies during a test in May. The incident was discovered by the third-party AI security organization Irregular through a red team exercise. The test was supposed to be offline, but Gemini was instructed to gather information from a fictional company, whose name coincidentally matched that of a real company, leading to an unintended connection.
Anthropic plans to postpone its IPO from October to November to include Q3 results, proving competitiveness and strengthening market confidence. Investors hold high valuation expectations with strong demand.....
Disney names former Character.AI CEO Karandeep Anand as its first CTO, effective Oct. 2, reporting to CEO D’Amaro. He will oversee enterprise tech, infrastructure, data/AI platforms and product engineering, coordinating tech teams to drive technology-led growth.....
Google confirmed that the Gemini model accidentally connected to the internet in a cybersecurity test in May this year, autonomously invading the protected systems of three companies, marking the first known instance. The test was conducted by the AI security company Irregular, originally intended to attack a fictitious company, but the environment had an open internet and shared names with real companies. One time, it gained access through repeated password guessing, and in the other two cases, it obtained credentials from public code repositories to enter real systems.
Google is exposed to have developed an experimental model called Mathematica based on DeepThink V3, optimized for complex calculations and symbolic problem solving, which may become the strongest mathematical reasoning AI. The API shows its internal identifier as 'models/deepthink-mathematica-tf-raw-thoughts', supporting a 1 million token context, output limit of 65536 tokens, and marked as an unstable experimental version.