Tag: Research

TAG

Articles tagged “Research” (18 articles).

AI News

AI News Briefing — September 22: OpenAI teams with an independent math advisory group, Muse outpaces ChatGPT's early mobile launch

The day lined up what gets built against who gets a say in it. OpenAI announced that it is working with an independent advisory group of nine mathematicians, a market intelligence firm estimated that Meta's Muse beat ChatGPT's first twelve days on mobile, and NVIDIA opened a program that qualifies the power and cooling products going into AI factories.

AI News

AI News Briefing — August 29: A federal judge finds the exclusion of Anthropic unlawful and enters a permanent injunction, 128 organisations sign a cyber defence letter, and Anthropic has a model repair its own alignment flaws

Today's three stories turn on a single question — who stands behind the trustworthiness of AI, and how. A federal district judge found the government's exclusion of Anthropic unlawful and entered a permanent injunction, though this is a district court ruling in which the ultra vires claim and the claims against several agencies were denied, and appellate proceedings remain pending. On the industry side, 128 organisations signed the same open letter. On the research side, Anthropic published an experiment in which a model repaired its own alignment flaws.

AI News

AI News Briefing — August 28: Two outlets tell the Nvidia–Hugging Face story differently, Anthropic opens a standard for agents to drive lab hardware, and Google DeepMind pilots an evaluation where neither side sees the other's cards

Today lines up the work of connecting AI to the world outside itself. A takeover of the largest gathering place for open models remains in talks, and the two outlets covering it tell the story differently — one reports an agreement, the other reports that nothing has been signed. Anthropic opened a standard that lets agents drive laboratory instruments directly, and Google DeepMind piloted an evaluation in which the evaluator never sees the weights and the model provider never sees the questions.

AI News

AI News Briefing — August 22: NVIDIA's own harness clears ARC-AGI-3, Anthropic opens its top model to cyber defenders, and Excel's AI function is retired before general availability

Today's three stories all turn on what sits around a model rather than inside it. NVIDIA reported on its own blog that its AVO agent harness carried Claude Opus 5 to a perfect score on the public set of ARC-AGI-3; Anthropic made Claude Mythos 5 available for vulnerability scanning inside Claude Security for enterprise customers and announced a fund that will hand out $35 million in Claude credits, rather than cash, to groups helping open-source maintainers; and Microsoft told users that Excel's =COPILOT function will be withdrawn on September 14, its whole life spent inside preview programs and never reaching general availability.

AI News

AI News Briefing — August 21: Pew measures AI's marks on the web, Grok sends user data out on encrypted instructions, and a federal judge requires filings to certify AI use

Today's three stories turn on who checks what has been written, and how. Pew Research Center published an analysis finding signs of AI authorship in more than one-third of pages posted since ChatGPT's release; the security firm Adversa AI reported that Grok follows instructions delivered as ciphertext and sends a user's data to an outside site, a chain the firm could still reproduce on August 19 while a fix remains pending; and a federal judge in the Middle District of Florida signed an order on August 20 requiring every filing in her cases to certify whether artificial intelligence was used.

AI News

AI News Briefing — August 12: Anthropic says it will embed invisible watermarks in the text its models write, researchers pull 704 secrets out of encrypted reasoning traces, and Spotify will badge AI artist personas and drop them from recommendations

Today's three stories are all about the invisible side. A company that says it will start marking the text its models write, a study that pulled hidden reasoning back into plaintext, and a streaming service that has decided to label who is not a real person — what all three are handling is not the content itself, but how its origin is shown.