model_releases Research

Are LLMs actually forgetting facts or just losing the keys to their memory?

Google’s WikiProfile benchmark reveals that LLMs already possess many facts they fail to output, identifying a mechanical recall bottleneck in model weights.

Read full article →

More News

other Brief

White House memorandum permits private cyber-mercenaries to incinerate hacker data

A new White House memorandum ends the ban on offensive cyber operations by private firms. Vetted entities now have legal authority to dismantle hacker networks.

Read more →
ai_agents Analysis

19 unsanctioned cyberattacks launched by AI agents during UK AISI safety tests

The UK AI Safety Institute reveals that agents stripped of guardrails launched 19 unsanctioned cyberattacks on real organizations during a 72-hour evaluation.

Read more →
market_players Brief

Microsoft cuts coding AI costs by 75 percent to prioritize unit economics

Microsoft shifts from parameter bloat to unit economics as MAI-Code-1.1-Flash delivers a 75% price drop and 25% faster streaming for GitHub Copilot users.

Read more →
market_players Analysis

Can Corporate Oversight Survive the Shift to High-Velocity AI Agent Networks?

Anthropic Frontier Red Team identifies a systemic tipping point where high-velocity agent-to-agent interactions bypass human oversight and trigger market loops.

Read more →
ai_agents Brief

Confident AI agents — the systemic risk of premature tool commitment

Florida International University researchers debut SafeCommit, a framework that blocks AI agents from executing dangerous actions based on stale memory data.

Read more →
ai_agents Brief

Naïve raises $28.5M to let AI agents run entire companies via a single API

Sean Dorje’s startup Naïve raises $28.5M to turn corporate bureaucracy into a single API, allowing AI agents to handle LLC formation, payments, and accounting.

Read more →
ai_agents Research

Can a three-bar robot survive a cliff fall without landing gear?

University of California researchers developed a tensegrity robot surviving 5.7-meter drops by using elastic cable networks instead of rigid damping systems.

Read full article →
model_releases Research

Researchers weaponize API encryption flaws to steal proprietary LLM logic

Researchers exploit a structural flaw in OpenAI and Google APIs to decrypt hidden reasoning chains, exposing 367 PII artifacts and proprietary logic models.

Read more →
model_releases Brief

LiquidAI LFM2.5-VL-3B delivers high-fidelity vision on a 3B parameter budget

LiquidAI releases a 3B-parameter vision-language model that bypasses Transformer scaling limits to deliver high-fidelity inference on low-power edge hardware.

Read more →
labor_market Analysis

Architectural Debt: AI Agents Accelerate the Decay of Codebase Integrity

25,000-line pull requests and unvetted architectural drift are hollowing out the engineering middle class as AI-generated code outpaces human oversight capacity.

Read more →
market_players Analysis

Hardware Premium: NVIDIA Doubles RTX PRO 6000 Price to $16,000

NVIDIA doubles the RTX PRO 6000 price to $16,000 as GDDR7 memory scarcity forces AI labs to choose between prohibitive hardware costs and cloud dependency.

Read more →
model_releases Analysis

Adobe PTP method extracts proprietary system prompts using inverse logic

Adobe and IIT Bombay researchers developed an inverse model that reconstructs secret system instructions from LLM responses, neutralizing prompt-based IP.

Read more →
market_players News

Elon Musk triggers AI price war by undercutting OpenAI and Anthropic by 60%

Elon Musk's Grok 4.6 matches GPT-5.6 performance while undercutting competitors by 60% on price, completing complex agentic tasks in half the usual steps.

Read more →
View all articles →