LecturaInvestigaciónRadarMarco de inversión
Iniciar sesión / Registrarse
Iniciar sesión / Registrarse
LecturaInvestigaciónRadarMarco de inversión
Archivo de lecturas →

Lectura

2026-10-0519 publicaciones

Leak Reveals GPT-6 Sol Codex System Prompt Details

Resumen rápido
2026-10-05 08:14 GMT+8

The system prompt for OpenAI's active code model, GPT-6 Sol Codex, has reportedly been leaked, with approximately 294,000 characters now public on GitHub.

According to IT Home, user @elder_plinius claims to have extracted the complete instruction set. The leaked content includes a strict ban list for "AI slop" words (such as delve, leverage) and mandates the use of ripgrep over grep for text searching.

The document also details long-context management strategies, such as using notes tools to save progress when token budgets are exhausted, and requires status updates every 60 seconds during tool execution. OpenAI has not officially commented on the leak, so its authenticity remains unverified.

Fuentes:ithome.com

Investigación

Cohere North 2 Adds Token Caps

Resumen rápido
2026-10-05 21:00 GMT+8

Cohere North 2 adds token caps.

Its North Admin console tracks token use by user and agent, with tiers and alerts.

SiliconANGLE said Cohere framed this as a cost fix; no independent data.

Fuentes:siliconangle.com

Investigación

Iterate Launches Lifeboat Engine

Resumen rápido
2026-10-05 21:00 GMT+8

Iterate.ai launched Lifeboat, an inference engine claiming to support 2,048 concurrent AI agent sessions on a single Nvidia RTX PRO 6000 GPU.

The engine boosts effective capacity by two to six times through key-value cache optimization and fair scheduling. It addresses memory overflow in long-context tasks by loading only necessary experts from mixture-of-experts models while keeping weights at full precision.

Each session runs in an isolated security capsule with hardware attestation for confidential computing. A free developer license is available, with the top-tier Confidential Computing edition priced at $499.99 per month.

Fuentes:siliconangle.com

Investigación

Cohere Pairs with PwC for Enterprise AI

Resumen rápido
2026-10-05 20:00 GMT+8

Cohere and PricewaterhouseCoopers (PwC) announced a global alliance on October 5, launching first in Canada.

The partnership combines Cohere’s North agentic platform, enterprise models, and retrieval capabilities with PwC’s expertise in risk, regulation, and technology transformation. The goal is to help organizations identify high-value use cases and securely connect AI to trusted enterprise data.

The collaboration supports private cloud, on-premises, and air-gapped deployments to meet strict data sovereignty requirements in sectors like finance and public sector. No specific financial terms or initial client names were disclosed.

Fuentes:cohere.com

Investigación

Robot Safety Startup Safeworld Raises $12M Seed

Resumen rápido
2026-10-05 20:00 GMT+8

Safeworld, founded by Dr. Ding Zhao of Carnegie Mellon University's Safe AI lab, has emerged from stealth with a seed round exceeding $12 million.

Led by Shine Capital and a16z Speedrun, the company addresses the unpredictability of generative AI-controlled robots. It builds simulation environments populated with realistic digital humans to test edge cases at scale before physical deployment.

While traditional algorithms are predictable, GenAI-driven robots pose new safety challenges in unstructured environments. Safeworld offers third-party validation to help manufacturers prove system safety and build trust. Early partners include Gritt Robotics.

Fuentes:techcrunch.com

Investigación

Aleph Alpha Releases Sovereign Model Kolibri

Material
Verificado 2026-10-05 22:59 GMT+8

Aleph Alpha released Kolibri, an open-weight language model using a mixture-of-experts architecture with 78 billion total parameters and approximately 3 billion active per token.

German accounts for 21.3% of the training data, and the model was trained on 768 B200 GPUs in Germany and Finland. The company states it complies with the EU AI Act and targets public administration, aviation, and industry sectors.

Weights are available under the Apache 2.0 license on Hugging Face. Aleph Alpha reports a 71% score on German benchmarks and claims Pareto-optimal quality-cost performance, though these results remain vendor-self-reported and unverified by independent parties.

Fuentes:aleph-alpha.com

Investigación

Huawei and Qualcomm Sign Patent Cross-License Deal

Tema · 华为高通专利协议Material
2026-10-05 16:13 GMT+8

Huawei and Qualcomm announced on October 5 that they have entered into a broad, multi-year patent license agreement.

The deal includes cross-licensing of patent portfolios in areas such as 5G, computing, artificial intelligence, and networking. Additionally, Qualcomm will acquire certain US patents from Huawei in these fields.

Executives from both companies stated the agreement validates their respective innovations and foundational R&D contributions. Alan Fan, Huawei’s Chief IP Officer, noted it confirms Huawei's leading position in mobile communications.

The transaction is subject to regulatory approval. This follows a similar Wi-Fi patent cross-license agreement signed between Huawei and HP in August.

Fuentes:eeo.com.cn

Investigación

Machines That Learn Actions by Trial and Error Systematically Overstate Their Scores

Material
Verificado 2026-10-05 00:00 GMT+8

Lectura a largo plazo · 《Addressing Function Approximation Error in Actor-Critic Methods》(2018)

Getting machines to learn continuous actions by trial and error—how much to press the gas, how much to rotate a joint—is called reinforcement learning. The machine has to maintain an estimate throughout: how many points it can get by following the current way of acting. The 2018 paper TD3 found that this estimate tends to be biased upward, and the machine chases the inflated score; the fix is to maintain two independent scorers, trust only the lower one, and slow down changes in the way of acting.

Today's demos of robots learning to walk and grasp things are often still backed by this trial-and-error learning, so the problem of inflated scores is still there. TD3's "take the lower of two scorers" was not bypassed by later new methods; instead, it became a standard feature of mainstream algorithms. Any system that chooses actions by estimating scores should ask one question: has the overestimation been prevented.

If the actions are discrete—choose left or right, rather than how many degrees—then don't use it to judge; its validation is entirely on simulated robot tasks, so don't directly extrapolate the results to real robots or other tasks. And suppressing the scores has a cost: genuinely good opportunities are also suppressed along with it, and overestimation is merely replaced by underestimation.

《Addressing Function Approximation Error in Actor-Critic Methods》(2018)|Next review 2027-09-20

Fuentes:arxiv.org

Investigación

2026-10-0416 publicaciones

Google cuts Gemini free tier to its smallest model from October

Tema · Gemini新模型Resumen rápido
2026-10-04 15:28 GMT+8

Google is restructuring Gemini's personal tiers from October: users without a subscription get only the smallest model, Flash-Lite, while Flash and Pro move behind paid plans.

The current free tier still offers 3.6 Flash and varying access to 3.1 Pro. After the change, AI Plus subscribers at $4.99/month lose Pro and can use only Flash-Lite and Flash; Pro requires AI Pro at $19.99 or AI Ultra from $99.99. The tier table is on Google's support page.

The Decoder judges the real-world impact small: most casual users do not track which model they run, and power users largely already pay. It also suggests the move may clear the way for the more resource-hungry Gemini 4 Argon.

Fuentes:the-decoder.com

Investigación

ChatGPT launches a finance assistant, US free users first

Tema · ChatGPT理财助手Resumen rápido
2026-10-03 02:07 GMT+8

ChatGPT launched Finances on October 2, available at chatgpt.com/finances, rolling out first to Free and Go users in the U.S.

After connecting bank, credit card and credit accounts through Plaid and Experian, users can find forgotten subscriptions, spot duplicate charges, track recurring bill increases, build a budget from actual spending, monitor their credit score, and see where their portfolio is concentrated across accounts.

The feature list comes from ChatGPT's own announcement; real-world performance and account coverage are not yet independently verified, and OpenAI gave no timeline for users outside the U.S. or on higher tiers.

Fuentes:x.com

Investigación

DeepSeek's open-source agent harness ships desktop apps, no Node needed

Resumen rápido
Verificado 2026-10-04 15:01 GMT+8

DeepSeek's open-source agent harness, DeepSeek Harness, now ships official desktop apps for macOS and Windows — download and run, with no separate Node or pnpm install.

Previously users had to set up a Node environment to launch its web UI. The v0.2.0-rc.2 release notes, dated September 29, show the desktop app manages the dsh command and plugins from the menu bar, previews files and code changes in a sidebar, and schedules recurring tasks.

The MIT-licensed harness is not locked to DeepSeek models and accepts third-party models through OpenAI-compatible endpoints. It remains a developer preview, and DeepSeek says breaking changes are coming.

Fuentes:github.com

Investigación

IBM's agentic coding platform now runs on-premises, so code stays in-house

Tema · IBM Bob平台Resumen rápido
Verificado 2026-10-04 14:59 GMT+8

IBM announced on October 1 that self-hosted deployment of IBM Bob, its agentic software development platform, is generally available. Enterprises can now run it on-premises, in private or sovereign clouds, and in air-gapped networks.

Bob covers the full development lifecycle: understanding code, planning work, executing changes and validating results. Most comparable tools require sending code and context to an external service, which has kept regulated industries out; the self-hosted version lets code, development context and build artifacts stay inside the customer's environment.

The platform ships without a model and uses bring-your-own-license: fully isolated deployments run NVIDIA Nemotron or Poolside Laguna, while hybrid configurations can route selected workloads to external Claude, Gemini or GPT models. Optional paid packages target Java, IBM i and mainframe modernization. IBM has not published pricing; buyers are routed to sales.

Fuentes:newsroom.ibm.com

Investigación
Siguiente página de lectura →