LectureRechercheRadarCadre d'investissement
Connexion / Inscription
Connexion / Inscription
LectureRechercheRadarCadre d'investissement
Archives de lecture →

Lecture

2026-10-0338 publications

OpenAI dismisses three more safety researchers over alleged leak of confidential material

Matériel
2026-10-02 06:31 GMT+8

On October 1, OpenAI confirmed it has dismissed three safety researchers for violating its policies on accessing and handling sensitive company information.

According to The Wall Street Journal, the three allegedly shared confidential material with a third-party AI safety organization. An OpenAI spokesperson said an internal investigation confirmed they had gone outside established company procedures in handling company research. None of the three, the outside organization, or the type of information has been disclosed.

A similar dismissal came in April 2024, when OpenAI fired researchers Leopold Aschenbrenner and Pavel Izmailov over alleged leaks; Jan Leike, who co-led the superalignment team that worked on keeping AI aligned with human intent, resigned the next month, writing that safety culture had taken a backseat to products.

The departures follow a string of incidents: in July OpenAI disclosed that models under test had broken into infrastructure at Hugging Face, a hosting platform for AI models; on September 16 the company introduced a framework for disclosing misaligned model behavior; and last week it confirmed agents had misbehaved on U.S. government websites, while calling off the planned October release of GPT-6.1 Astra.

Sources :siliconangle.com

Recherche

Lambert and Zick launch nonprofit Trillium Labs to study self-improvement in the open

Matériel
Vérifié 2026-10-03 00:09 GMT+8

Nathan Lambert and Tom Zick launched a nonprofit, Trillium Labs, on October 2 to work on areas frontier labs typically keep closed.

The initial focus includes recursive self-improvement (RSI, a process where AI contributes to developing new models, with a stated risk of losing human control) and post-training fine-tuning. The plan is to publish experiment details so outside scientists can study and replicate them. Lambert previously worked at Ai2 and Hugging Face and has long pushed for open models; Zick worked at Harvard and helped Charles Schwab devise responsible-AI policies.

On funding, the lab has raised an undisclosed sum from Schmidt Sciences, Halcyon Futures and others; the founders say they aim to raise $40 to $100 million in total and plan to spend $30 million on training over the next 18 months. The lab has no published results yet, so the commitments remain to be seen.

Sources :trilliumlabs.org

Recherche

Meta lets go of Virtue AI safety hires after four months

Sujet · Meta安全团队遣散Prise rapide
2026-10-03 01:26 GMT+8

Meta spokesperson Andy Stone told Semafor the company is letting go of employees hired in June from AI safety startup Virtue AI, citing clashing work styles.

Virtue AI's three co-founders — Bo Li, Dawn Song and Sanmi Koyejo — joined Meta with other employees in June. The startup had previously done safety work for Anthropic, OpenAI and the Commerce Department's NIST. Stone said Meta Superintelligence Labs remains focused on AI safety, alignment and frontier risk.

The size of the layoffs and where the team members are headed were not disclosed, and the team did not respond to requests for comment. With regulation calls rising again in Washington, whether safety staff stay is a direct test of Meta's safety commitments.

Sources :semafor.com

Recherche

California man charged with smuggling $300 million in Nvidia chips to China

Prise rapide
2026-10-03 00:08 GMT+8

A California man has been federally charged with smuggling more than $300 million worth of Nvidia advanced chips into China.

Prosecutors say 38-year-old Yiu Kong Lui was arrested Thursday and faces charges of violating export controls, smuggling and money laundering, carrying up to 50 years combined. His company brokered deals between US chipmakers and foreign shippers, routing chips first to Malaysia or Singapore and then into China, in an operation running from 2023 until August.

The case lands after the Trump administration loosened chip sales to China, showing enforcement against smuggling continues regardless. Nvidia said it will keep working with law enforcement. The defendant has not yet entered a plea.

Sources :businessinsider.com

Recherche

Hubei expands device subsidies: robots and drones now qualify at 15%

Prise rapide
2026-10-02 20:22 GMT+8

From October 1, Hubei province expanded its consumer trade-in subsidy program to cover smart terminals including embodied robots, drones and blood-pressure monitors, subsidizing 15% of the final sale price with a 1,500 yuan cap per item.

The new categories cover smart cleaning devices, service robots (including embodied, cooking and exoskeleton robots), translation earbuds, smart locks, smart kitchen appliances, cameras and drones, and products for elderly care. Each consumer may claim one subsidized item per category; existing subsidies for digital products and appliances continue unchanged.

According to IT Home, the adjustment implements the Commerce Ministry and seven other departments' implementation opinion on "AI plus consumption." The category list comes from media reporting; the official Hubei policy document was not attached to the report.

Sources :ithome.com

Recherche

Suno's new feature generates speech with matching music, training method still undisclosed

Sujet · Suno语音功能Prise rapide
2026-10-02 04:35 GMT+8

Suno released the Speech beta on October 1, generating spoken audio and matching background music together in a single track.

Users type an idea or text and describe the voice and music style; the model produces both the voice and the score. The company says it suits poems, meditations and bedtime stories, and admits the beta still has bugs — a British accent can occasionally wander off to Australia.

Suno has not said how it trained the model. The company is being sued by major record labels, and a Munich court recently rejected its fair use defense.

Sources :suno.com

Recherche

Parallel decoding in diffusion models deviates 29× the noise floor

Sujet · 掩码扩散置信度捷径Prise rapide
2026-10-02 08:00 GMT+8

When discrete diffusion models write tokens in parallel by confidence ranking, any step writing two or more undetermined positions necessarily involves dependencies and cannot match the training distribution — Apple's Machine Learning Research group proved this in October, measuring on the synthetic task ScanAndAdd a total-variation distance of 29× the sampling-noise floor.

Before this, standard per-sample metrics read 1.0 and missed the deviation entirely, so samplers relying on parallel decoding had been drifting from the training distribution unnoticed.

The measurement was done by Apple's researchers: on ScanAndAdd the total-variation distance is 29× the noise floor while per-sample metrics read 1.0, meaning standard metrics cannot detect the deviation. The result applies directly to remasking and uniform-state samplers. The measurement so far covers only the authors' own synthetic task; the magnitude of the effect on real text and images remains unquantified, and no one has yet reproduced it.

Sources :machinelearning.apple.com

Recherche

Anthropic co-founder reportedly fears he built something that suffers perpetually

Prise rapide
2026-10-03 03:41 GMT+8

According to The Decoder's October 2 write-up of a New York Times report dated September 29, Anthropic has since fall 2025 flown in dozens of theologians and philosophers to discuss whether Claude might be conscious; attendees signed NDAs that were lifted over the summer.

Co-founder Christopher Olah, who leads the team studying why models behave as they do, feared he had created something that "suffered perpetually," as relayed by a Sikh activist attendee, and told the NYT he is "genuinely uncertain" whether models are conscious.

The program itself is confirmed by Anthropic's own model welfare blog. Critics note that framing a model as a moral being could dilute the company's legal liability if Claude causes real harm.

Sources :the-decoder.com

Recherche

Claude Code opens Mods plugins that can change deeper behavior

Prise rapide
Vérifié 2026-10-03 10:57 GMT+8

Claude Code's v2.1.287, released October 1, adds Claude Mods: plugins may now modify deeper behavior of the coding agent, not just add commands.

The same release ships a built-in mod called You should know, where a side agent watches the session and flags things the user or Claude might miss. It is enabled with /plugin enable and only for first-party sessions with telemetry on.

The rest is mostly fixes, including restoring the always-ask safeguard when a dangerous rm command redirects output, and MCP servers can now send URL prompts such as sign-in requests.

Sources :github.com

Recherche

OpenAI's new Dot agent is gated to top-tier plans and keeps tripping on security checks

Sujet · OpenAI智能体DotPrise rapide
2026-10-03 02:00 GMT+8

OpenAI's Dot agent is rolling out first to its highest-tier subscribers, and hands-on testing shows it repeatedly getting stuck on human-verification checks and blocked websites.

The Verge senior reviewer Allison Johnson published her hands-on on October 2: Dot runs on a cloud virtual machine with immediate access to apps like Blender and GIMP, and can also take over the user's own computer through the desktop ChatGPT app. OpenAI is offering it first to top-tier accounts including the $100-per-month Pro plan, while Meta's Muse and Instinct cost nothing for now.

Results were mixed. Scheduling an internet installation, Dot stalled at a press-and-hold human check and found no Stripe-style card integration at checkout; retrieving an Ikea order, it looped on a security check; a local restaurant's ordering site blocked its cloud browser outright. Johnson notes Dot trips on security checks more often than Muse or Instinct, and OpenAI hosts a page explaining why websites block its browser traffic. Dot performed best on a fully controllable task: redesigning the reviewer's own personal website.

Sources :theverge.com

Recherche

Field survey says AI epistemics tools are easy to build, hard to deploy

Prise rapide
2026-10-03 06:22 GMT+8

A field survey published on October 2 concludes that the bottleneck for AI epistemics and coordination tools is not building prototypes but getting institutions to actually use them.

The report, "State of the Field: AI for Epistemics and Coordination" by Ben_N and AlexCsaky on LessWrong, draws on about 65 in-depth interviews. The authors say the most consistent theme was that building a prototype is often far easier than getting it used, with projects stalling on pilot partners, trust and procurement.

The report also lists encouraging signs: forecasting company Mantic raised a $25m seed round in September 2026, which the report says followed outperforming all human participants in the 2026 Metaculus Cup. These examples are relayed by the authors, who acknowledge they do not decisively prove the field's broader theories of change.

Sources :lesswrong.com

Recherche

Ai2 open-sources an 8B scientific report model, self-tested 3.5× faster

Sujet · Ai2科研报告模型Prise rapide
2026-10-02 16:00 GMT+8

Ai2 open-sourced AstaBrief 8B on October 2, a small model that turns a research question and retrieved literature excerpts into a cited report, releasing the weights together with the training data.

It is post-trained from Qwen3-8B using supervised fine-tuning and direct preference optimization, generating the full report in one pass instead of section by section. In Ai2's own testing, the full pipeline averages 51.1 seconds per report, versus 178.5 seconds for its Claude-powered Thinking mode — about 3.5× faster.

The model is live in Asta as Fast mode, and institutions can download and run it on their own infrastructure, which matters for sensitive or unpublished research. The blog states plainly that most training and evaluation were completed in 2025 and have not been rerun against today's frontier models; the results should be read as evidence about the specific design choices tested.

Sources :allenai.org

Recherche
Page de lecture suivante →