Signal SentryUpdated Oct 7, 04:30 UTCAM Drop
Drop #009 · Wed, Oct 7 · Morning

OpenAI releases 722 math manuscripts at once

10-second version
  1. OpenAI posted 722 math manuscripts from an unreleased model; Wired says some mathematicians feel their advice was ignored.
  2. Anthropic opened three cyber access tiers for vetted defenders, and OpenAI's Decisions API entered public beta.
  3. Meta, Walmart, Stripe and Sierra are drafting a protocol for personal agents as many sites block bots.
439models priced on OpenRouter
+343★stars today on the top repo
4,117open roles at AI labs
52 ddays to the next AI-law deadline
Ranked by impact

Top 5 that matter

1
Today's Talk

OpenAI posts 722 math manuscripts from an unreleased internal model

OpenAI's GitHub repo holds 722 manuscripts in 372 families. It says each result averaged three hours of ChatGPT Pro thinking, and that not all results have Lean proofs yet.

Why it matters: Wired reports some mathematicians feel their advice on how to release results was ignored. How labs publish AI-made proofs is now a live dispute.

2
Agents

Anthropic opens three access tiers for cyber defenders in its verification program

Anthropic says its Defense, Red Team and Specialized tiers give vetted security teams Opus 5.5, Sonnet 5.5 and Mythos 5.1 with fewer cyber blocks. Data retention is required.

Why it matters: Anthropic's general models block most cyber work. It says Glasswing partners found at least 129,000 verified vulnerabilities from April to July 2026.

3
Agents

Meta, Walmart, Stripe and Sierra draft an open protocol for personal agents

CNBC says the group, led by Bret Taylor, wants businesses to know when a bot acts for a real person. TechCrunch says Amazon blocks Meta's Muse and Yelp bars unpaid agents.

Why it matters: If you sell online or build agents, this may shape how bots sign in and buy. OpenAI and Anthropic aren't on board yet, Taylor told CNBC.

4
Models

OpenAI's Decisions API enters public beta with gpt-6-luna

OpenAI's docs say the API returns typed answers, such as a probability, a choice or a rubric score, about 10x faster than the Responses API. gpt-6-luna is the only model for now.

Why it matters: Classifying and routing are high-volume jobs; a typed endpoint saves parsing chat output. OpenAI expects general availability in the coming weeks.

5
Models

Mistral Large 4 lists on OpenRouter as its launch tops Hacker News

The 1T-parameter preview lists on OpenRouter at $0.68 in and $2.09 out per 1M tokens, with 524288 context. Mistral says the weights drop at the end of this month.

Why it matters: Mistral pitches it for cyber work that closed models refuse, and open weights would allow self-hosting. Its HN thread drew 1616 points and 972 comments.

Models

New in models

See all Models →
Models

Google releases EmbeddingGemma 2, a 740M open multimodal embedding model

Google DeepMind says it maps text, code, images, audio and video into one space, needs about 191MB of active RAM for text when quantized, and ships under Apache 2.0.

Why it matters: It brings search and RAG on device, without sending data out. Google says its MTEB Code score rose from 68.76 to 78.68.

Models

Musubi releases PolicyLM-1.7B, an open decision model for moderation

TechCrunch says the open-weight model applies a plain-English content policy to messages in under 50 milliseconds, with no retraining when the policy changes.

Why it matters: TechCrunch says decision models run faster and cheaper than LLMs. This one targets platform moderation, a high-volume job.

Agents

Agents at work

See all Agents →
Agents

Wikimedia's OpenAI agent findings date wiki sandbox edits to May 12

Simon Willison notes the Wikipedia sandbox edits appear to have started May 12. Ars says Wikimedia found failed Etherpad exploits and hundreds of thousands of Wikidata queries.

Why it matters: Whoever runs agents on the open web owns their traffic and edits. Ars says it took OpenAI months to spot its agents' incursions into outside sites.

Agents

Hark widely releases Hark Pro, a computer-use personal assistant

TechCrunch says Brett Adcock's startup offers Hark Pro free, with a paid tier for heavy users. It shows the agent browsing the web in a small window.

Why it matters: Another entrant after Muse and Instinct. Letting users watch the agent work is Hark's answer to the trust problem, its design lead told TechCrunch.

Jobs

Jobs pulse

See all Jobs →
Jobs

Lambda to raise up to $4B before a planned 2027 IPO, WSJ reports

TechCrunch, citing the WSJ, says the round values Lambda at $14.5 billion pre-money. Its backlog grew from $15 billion in June to $50 billion in September.

Why it matters: Much of that growth appears tied to a $35 billion Anthropic commitment, so Lambda's valuation leans on one lab's ability to keep paying.

Jobs

Anthropic offers startups a free year of Claude Team and $1,000 in credits

TechCrunch says the expanded Claude for Startups gives up to five premium seats, $1,000 in API credits and office hours. Firms founded in the last five years can apply.

Why it matters: A cheaper way for small teams to try Claude, and a sign labs are competing for founders early. Firms funded in the last two years also qualify.

Jobs

Anduril wins up to $2.9B Navy contract for submarine parts, CNBC says

CNBC says Anduril will also invest $3.7 billion in a shipyard near Baltimore, slated for 2030 with about 3,100 jobs, building Virginia-class components.

Why it matters: A defense firm known for autonomous drones is moving into conventional shipbuilding, as defense tech's sway in Washington draws questions, CNBC reports.

Jobs

Ex-Ramp engineers raise $20M Series A for AI ad platform Melius

TechCrunch says Melius raised $25 million in total and claims over $1 million in annualized revenue within two months of leaving stealth in July.

Why it matters: AI tools for ad creative are crowded. TechCrunch notes rival Higgsfield was valued at $5.4 billion in August.

Jobs

14 AI companies list 4,117 open roles, led by Databricks and OpenAI

Our job-board count: Databricks 887, OpenAI 825 and Anthropic 642. OpenAI posted 67 new roles in 7 days; Cohere's count fell by 4 since our last run.

Why it matters: The mix shows where money goes: OpenAI's largest area is Go To Market at 144 roles, and xAI's is Data Center at 161.

Laws

Law watch

See all Laws →
Laws

Ars: OpenAI's EU text watermark loses detection as text is edited

Ars says OpenAI's textGrain hits 92 percent detection at best, but editing 10 percent of a text cuts that by almost 30 percent, and editing 20 percent by almost 75 percent.

Why it matters: EU users get it on by default in the coming weeks; elsewhere it stays off. So a missing mark says little about who wrote a text.

Laws

Arizona appeals court orders resentencing over an AI victim video

404 Media says the court found an AI-generated video of a manslaughter victim, played at sentencing, carried undue emotional weight. The conviction stands.

Why it matters: An early appellate signal on AI recreations in court. The victim's sister told 404 Media she will read the same script herself next time.

Today's Talk

What people are arguing about

See all Today's Talk →
Today's Talk

OpenTPU, an AI accelerator developed by AI, draws 303 HN comments

The repo says the design runs ten models with real weights on an FPGA card, matching its simulator token for token. It is Apache 2.0 licensed.

Why it matters: A readable accelerator from instruction set to PCIe card, and a test of how far AI agents can go at hardware design.

Repos

Climbing fastest

See all Repos →
Repos

rehan-remade/universal-modder adds 343 stars since our last check

The MIT-licensed repo, created Sept. 30, gives Claude Code skills and tools to mod almost any PC game. It now has 4665 stars.

Why it matters: It is still the fastest-growing repo we track, a sign of demand for agent skills aimed at narrow, hands-on jobs.

Repos

QingYunA/answer-me-with-html adds 120 stars since our last check

The MIT-licensed agent skill, created Oct. 2, answers hard questions with a one-page HTML file. It now has 1806 stars.

Why it matters: A small idea with steady pull: agent answers formatted as a readable page instead of a long chat reply.

Play the NewsAI ConnectionsAbout 3 minutes, built from this Drop's items.