LesenForschungRadarAnlageframework
Anmelden / Registrieren
Anmelden / Registrieren
LesenForschungRadarAnlageframework
Lesearchiv →

Lesen

2026-09-2366 Beiträge

Anthropic ships its new flagship 20% cheaper, performance claims still self-tested

Wesentlich
2026-09-23 00:30 GMT+8

Anthropic released Opus 5.5 on September 22, cutting output token pricing to $20 per million tokens from $25 for the previous model, a 20% drop.

The company calls it the strongest-performing model it has tested and says it outpaces its own larger Fable model on many benchmarks. Those results are all vendor-run; METR and other outside groups did pre-release safety evaluation only, and no independent performance verification exists yet.

Anthropic says the model is comparable to Mythos in biology and cybersecurity capabilities, so it carries the same usage safeguards as Fable. Sonnet 5.5 and Haiku 5.5 are promised in the coming weeks; whether the price cut extends to mid-tier models is what buyers should watch.

Quellen:techcrunch.com

Forschung

OpenAI's new Sol and Luna ship at half price, reliability claims still self-tested

Wesentlich
2026-09-23 02:00 GMT+8

OpenAI released GPT-6 Sol and Luna on September 22, pricing the API at half the cost of the 5.6-series equivalents, which the company attributes to improvements in caching and inference.

Sol targets complex tasks like coding; Luna handles high-volume clerical work such as summarizing and extraction. Both are live in ChatGPT Work, Codex and the API, and Luna will also reach Free and Go users.

The reliability claim deserves a discount: the reported halving of mistakes comes from OpenAI's internal evaluation based on user-flagged conversations, not third-party testing. Anthropic shipped Opus 5.5 just 90 minutes earlier, and the pricing race between the two is clearly deliberate.

Quellen:techcrunch.com

Forschung

OpenAI halves new model prices, entering DeepSeek's low-price territory

Wesentlich
2026-09-23 18:21 GMT+8

OpenAI released GPT-6 Sol and Luna with API prices roughly half the previous generation's; Luna costs $0.10 per million input tokens and $0.50 per million output tokens.

Per ifanr's report, Luna's standard-request pricing now sits in the range of DeepSeek V4.1 Flash's cache-miss rates (1 yuan input, 4 yuan output per million tokens), though DeepSeek's off-peak cache-hit input at 0.02 yuan remains far below Luna.

Performance figures are OpenAI's own claims; third-party Artificial Analysis estimates both models' intelligence is roughly flat versus GPT-5.6, with the main change being cost.

Quellen:ifanr.com

Forschung

Same AI performance now costs half as much each quarter; data-center revenue assumptions need a rethink

Wesentlich
2026-09-23 19:17 GMT+8

Epoch AI estimates that over the past three years, the cost of reaching a given level of AI performance has fallen about 47% per quarter on average — roughly 13-fold a year.

The report's example: in January 2025, o3 needed about $0.30 per question to score 75% on GPQA Diamond; in mid-2026, GPT-5.6 Luna hit the same score for $0.0004 — a roughly 725-fold drop in under 18 months. Math benchmarks fell fastest (50–52% per quarter), game-based puzzles slower (39–43%).

This is the authors' own analysis, and the report names its limits: benchmarks may be gamed, and it assumes users always pick the cheapest capable model, so the numbers should not be read as exact. Even at these prices, the report notes total spending could stay high as usage expands.

Quellen:marginalrevolution.com

Forschung

Alibaba's Qwen cuts audio API prices, ASR down as much as 95%

Wesentlich
2026-09-23 20:31 GMT+8

Alibaba's Qwen team released Qwen-Audio-3.1, a lineup of five audio models, and cut its voice API prices: TTS drops about 70 percent, Realtime roughly 85 percent, and ASR up to 95 percent.

The series covers speech recognition, text-to-speech and real-time interaction. Qwen says the ASR model improves multilingual and dialect recognition and cleans up filler words, TTS-Next generates voice and sound effects in a single diffusion pass, and the real-time model supports simultaneous speaking and listening with instant interruption.

All capability claims are Qwen's own; per The Decoder, they come from the official blog and an X announcement, with no independent evaluation yet. Cost assumptions for voice applications can be repriced at the new rates.

Quellen:the-decoder.com

Forschung

Muse agent zero-day is patched, but the design flaws remain

Wesentlich
2026-09-23 20:54 GMT+8

Meta's Muse assistant, launched only weeks ago, shipped with a zero-day: any local app or terminal command could rewrite the transcription server address and capture the account token, taking full control of the agent. Meta released a hotfix roughly 12 hours after disclosure.

The discoverer, macOS security expert Patrick Wardle, says attackers need not write full malware — the agent's own privileges suffice to write files or snap pictures with little or no indication. The root causes are design choices: cloud-based transcription, and any local process being able to control undocumented settings. The patch closes the hole; those decisions stand.

A separate fact: about 12 hours before disclosure, Amazon began blocking Muse from shopping on its site, calling it an unauthorized AI agent that violates its Conditions of Use, and asked Meta to remove Amazon from the experience. The exploit details are Wardle's own proof-of-concept, not independently reproduced; he plans to present them at a security conference in November.

Quellen:wired.com

Forschung

Meta's Muse Makes Calls for You — Sometimes a Human Is Making Them

Wesentlich
2026-09-23 00:33 GMT+8

Internal Meta posts seen by 404 Media confirm that the Muse feature, which calls businesses on a user's behalf, includes a human agent layer during testing, with some calls placed by trained human agents.

On September 16 Meta executives publicly promoted Muse's outbound calling. The internal posts state that Muse can hand a request to a trained human agent who places the call. Employees questioned users unknowingly handing call requests to people, and one warned of severe negative coverage.

A Meta spokesperson said this is internal dogfooding and that proper disclosures will come before any public release. How often calls are routed to humans, and under what conditions, remains undisclosed.

Quellen:404media.co

Forschung

Meta's Muse agent hits 500,000 users in week one, and Meta admits it is heavily inspired by OpenClaw

Wesentlich
2026-09-23 22:42 GMT+8

Meta's personal AI agent Muse drew more than 500,000 users in the week after its September 8 launch, including over 250,000 daily active users, according to internal data reported by The Information. The app has reached number one in Apple's App Store.

Nat Friedman, head of product at Meta Superintelligence Labs, wrote on X that Muse is "definitely heavily inspired as a product by OpenClaw" but was built from scratch; users had noticed nearly identical file names and contents, and Friedman replied that OpenClaw's creator "got those things exactly right." The user figures are Meta internal data relayed by media, not independently verified.

According to The Information, OpenAI has discussed building its own personal assistant in response. The real test is whether users will hand personal accounts and sensitive data to agents like this.

Quellen:the-decoder.com

Forschung

Meta adds over $200 billion in market value in two weeks as Wall Street warms to Muse

Wesentlich
2026-09-22 22:11 GMT+8

Two weeks after launch, Meta shares are up more than 20% cumulatively, adding over $200 billion in market value and hitting a seven-month high.

Per a Reuters report (relayed via IT Home), Apptopia data shows Muse reached 2.8 million downloads in its first 12 days, available only in the US and Canada; on a like-for-like basis with ChatGPT's first 12 days, Muse logged 1.8 million downloads versus ChatGPT's 1.3 million. LSEG data shows 57 of 64 covering brokers rate Meta a buy or better.

Jefferies' $10.8 billion annualized revenue figure is a scenario, not a forecast: it requires 1 billion users by end-2027 with at least 3% converting to paid. Downloads are not retention, and paid conversion is unproven.

Quellen:ithome.com

Forschung

Nscale's IPO prospectus never names the customer behind 73% of revenue

Wesentlich
2026-09-23 22:44 GMT+8

Nscale's 192-page US IPO prospectus does not name its largest customer, Bytedance, which accounted for 73 percent of the company's $33 million in 2025 revenue.

According to the Financial Times, relayed by The Decoder, the filing mentions only Bytedance's Singapore subsidiary, Spring, in an appendix. In May 2025, Spring contracted to use 2,304 Nvidia B200 chips in Norway, and that contract backed a $105 million loan from Macquarie.

Nscale expects its largest customer's share of revenue to fall below 20 percent this year. The disclosure picture comes from press reporting; the S-1 itself is not in evidence, so readers should check the original filing.

Quellen:the-decoder.com

Forschung

Hackers claim they took data on all FBI employees; FBI has not confirmed

Wesentlich
2026-09-23 00:48 GMT+8

The hacking group ShinyHunters claims it breached FBI-related systems and stole data on "all FBI employees and applicants." The FBI has not responded to a request for comment.

The group gave 404 Media a sample covering about 5,000 FBI employees, including addresses, phone numbers, dates of birth and, in some cases, spouse details. 404 Media ran some sample phone numbers through OSINT Industries and the names did correspond. The hackers say they used a zero-day in Oracle PeopleSoft, then reached AWS GovCloud servers and exfiltrated roughly two to three terabytes; they also defaced the FBI jobs site on Tuesday, which currently shows the applicant portal as unavailable.

The boundary: the core facts so far are the hackers' own claims, the sample was only checked for name-phone correspondence, and the "all employees" scale is unconfirmed by the FBI. The group says the motive is not financial; watch what it does next.

Quellen:404media.co

Forschung

Cognex to buy RealSense for about $600M, betting on robot 3D vision

Wesentlich
2026-09-23 07:42 GMT+8

Cognex said it intends to acquire RealSense, the depth-camera maker spun out of Intel 14 months ago, in a deal valued at around $600 million.

About $500 million of the consideration is cash, with the rest made up of a three-year retention program worth $56.5 million and roughly $50 million in restricted stock units. Closing is expected in the fourth quarter, subject to regulatory clearance.

RealSense raised $50 million when Intel spun it out, and its depth cameras let robots and drones navigate and avoid obstacles in 3D. The company says it expects $80 million to $90 million in sales this year, up more than 50%, and has been profitable for the last two quarters; these operating figures are not independently audited.

RealSense's facial authentication business is excluded and will be spun off as a separate company before closing. Intel still holds a 20% stake and one board seat. Cognex CEO Matt Moschner said the acquisition extends Cognex into robotic perception, toward a full-stack visual intelligence platform for physical AI.

Quellen:siliconangle.com

Forschung
Nächste Leseseite →