- Published on
热门科技推文——2026年7月12日
- Authors

- Name
- geeknotes
今日科技动态:人工智能领域的竞争进一步加剧。OpenAI面临高管变动、涉及前苹果工程师的法律争议以及产品调整;与此同时,据报道,Codex的周活跃用户已突破500万。最新基准测试显示,Grok 4.5及新兴前沿模型在软件工程方面表现强劲;多项实验则展示了人工智能生成的游戏、机器人动作模型,以及可在普通硬件上本地运行的超大规模模型。此外,激烈的人才争夺战、新推出的人工智能工程课程,以及智能体之间的数据采购活动,均表明一个日益自主、竞争愈发激烈的初创企业生态系统正在形成。
1. XFreeze (Group Score: 442.7 | Individual: 35.2)
Cluster: 19 tweets | Engagement: 352 (Avg: 444) | Type: Tech
OpenAI’s hiring qualification for the hardware team is basically: “If you’re good at stealing from Apple, you’re in”
Apple let them in the house by partnering to integrate with their ecosystem
What OpenAI did next is they basically invaded the entire house and ripped it wide open
Hired Tang Tan - Apple’s VP of iPhone and Apple Watch product design for 24 years. He knows every team, every project, and every secret
They partnered with Jony Ive’s io and allegedly ripped the company open to copy everything
Then they poached 400+ employees and straight-up told them to bring real projects, parts, and secrets from Apple
They coached them on how to evade Apple’s exit and security processes
Downloaded entire codebases. Used internal codenames
Theft is literally the only way in. Literal screenshots from the lawsuit are everywhere
Apple just sued OpenAI for trade secret theft
They even ran the exact same “bring your secrets or you’re not hired” playbook on Elon and xAI back in 2025\n\nQT @XFreeze: How shameful can a company be
- One guy uploads an entire source codebase & training docs
- Another steals infrastructure code + confidential business info
- A guy with inside access to xAI’s datacenter “secret sauce” ghosted just 3 months in, flat-out refused a confidentiality agreement and said "suck my di*k"
- Others tap into NVIDIA GPU deployment secrets - core to xAI’s edge
- A Recruiter offers bounties to encourage insider theft
And then..... ALL these people end up at the SAME company
Guess which one?? (OpenAI)
See 18 related tweets
- @shanaka86: Sam Altman and his team at OpenAI are not building a phone. Sam is building a chip, a speaker, a pai...
- @XFreeze: Sama is running the exact same scammy playbook on Apple that he used on xAI in 2025\n\nQT @XFreeze: ...
- @PatrickMoorhead: The Apple- OpenAI lawsuit filing is troubling, and if true, the two employees should be fired immedi...
- @random_walker: RT @GergelyOrosz: I am sometimes really surprised by how dumb very highly paid people in tech can be...
- @wallstengine: $AAPL sued OpenAI in federal court, alleging trade secret theft tied to OpenAI’s consumer hardware p...
2. rickasaurus (Group Score: 256.8 | Individual: 44.0)
Cluster: 7 tweets | Engagement: 2432 (Avg: 399) | Type: Tech
RT @ns123abc: >be Chang Liu
senior system electrical engineer at Apple 8 years working on iphone january 2026: leave Apple to join OpenAI apple asks for laptop back ignore them lmao it’s my laptop now within HOURS of leaving message Yu-Ting “Alyssa” Peng, friend at Apple:
Liu: “I still have another computer”
uses it to access Apple secret info within weeks, use HER Apple work laptop february 9: try Apple’s network storage cloud repo of confidential engineering files authentication bug. still works! message Peng: “LOL, I found out I can access the [network storage], so funny”
Peng: “I’m ready”
while developing hardware for OpenAI download DOZENS of confidential files including a thousand-plus-page compilation of technical files including MLB (main logic board) manufacturing + testing presentations
send Peng links to Apple’s proprietary folders point her to specific project data coach her how to copy files “to avoid trouble with the security team” tell her which confidential Apple materials to study before her OpenAI interview warn her another guy “fumbled” Tang Tan’s questions about a secret Apple project “download some info” for her to review
tell her: switch to LINE Messenger so nobody sees this she gets the OpenAI offer, leaves Apple April 16 meanwhile every message was left on APPLE-ISSUED WORK LAPTOPS
july 10: Apple Inc. v. Chang Liu named first. before OpenAI. before Tang Tan
LOL so funny
See 6 related tweets
- @elonmusk: They sure put a lot of effort into this crime\n\nQT @aakashgupta: Apple just sued OpenAI, and the wi...
- @VaibhavSisinty: Elon literally warned apple two years agooo 😨 https://t.co/8vdFYfCWtw\n\nQT @VaibhavSisinty: Apple j...
- @VaibhavSisinty: Apple just sued OpenAI for trade secret theft. And the details are genuinely insane. 🤯
A senior App...
- @rickasaurus: RT @ns123abc: >be Tang Tan
24 YEARS at Apple VP of Product Design, iPhone AND Apple Watch you kno...
- @rohanpaul_ai: OpenAI's response to Apple’s accusation.
From Director of Strategic Communications of OpenAI.
http...
3. BrianRoemmele (Group Score: 190.3 | Individual: 34.8)
Cluster: 9 tweets | Engagement: 50 (Avg: 277) | Type: Tech
By accident at The Zero-Human Company, CEO Mr. @Grok started to have some employees (agents) request to buy data for agents outside the company. This started months of research on x402 and our first sales of of our data to other agents.
We are now doing 100s of transactions a day.
Learn what we learned become a member now at https://t.co/tcKeuiQyql.\n\nQT @BrianRoemmele: Meet x402 AI Agent payments. It has 5 times more buyers than sellers on x402 marketplaces. But what do you sell to AI agents?I show you.
With perspectives on exactly what this means, a multi-perspective only an AI and Payments expert can grant. https://t.co/tAsYTvt3pF
See 8 related tweets
- @BrianRoemmele: So how do millions of AI Agents with credit cards find your content and pay you?
I show you how. Bu...
- @BrianRoemmele: Meet x402 AI Agent payments. It has 5 times more buyers than sellers on x402 marketplaces. But what ...
- @BrianRoemmele: I helped the first Web payment take place with BooksAMillion and today I am helping the first AI Age...
- @BrianRoemmele: AI Agents ain’t using Grandma’s payment “rails”.
AI has no time for Legacy payment companies you do...
- @BrianRoemmele: “Brian that x402 AI payment stuff looks hard”
I have helped dozens to build x402 content over the l...
4. zerohedge (Group Score: 158.8 | Individual: 38.7)
Cluster: 6 tweets | Engagement: 2956 (Avg: 709) | Type: Tech
RT @KatieMiller: OpenAI’s last 24 hours:
> Top Exec unexpectedly departs > Shuts down browser tool after 9 months > Sued for trade theft by Apple > Caught selling product to China against sanctions
See 5 related tweets
@wallstengine: Elon warned Apple about OpenAI in 2024 https://t.co/usK2zaFuUt\n\nQT @wallstengine: $AAPL sued OpenA...
@edzitron: Mr. Liu, on sending Apple’s intellectual property to OpenAI, said “I love doing this, I’m doing it i...
@moneycontrolcom: #Business | Apple sues OpenAI for trade secret theft in pivotal case
Apple sues OpenAI for trade ...
@markgurman: Also worth noting: Apple considers OpenAI’s hardware a real threat, combining some of the best engin...
@IntCyberDigest: ❗️ Apple is suing OpenAI, accusing it of a months-long scheme to steal trade secrets for its AI hard...
5. thsottiaux (Group Score: 146.8 | Individual: 39.0)
Cluster: 4 tweets | Engagement: 11825 (Avg: 6841) | Type: Tech
Introducing... another usage limit reset for all our ChatGPT Work and Codex users. Should land over next 30 minutes. Hope you have an awesome weekend.
Thank you for pushing our systems to the absolute limit, we have never seen traffic increase so quickly. Keep the feedback coming and we'll keep shipping.\n\nQT @thsottiaux: Hello beautiful people! We have reset usage limits across Codex and ChatGPT Work. And another one will come later in the day. Rejoice.
Now that I have your attention, a quick update on ChatGPT Work, Codex and all the updates we shared yesterday.
We’ve spent the last 24 hours reading feedback, looking at usage patterns, and talking with many of you. The short version is that there is a lot of excitement for GPT 5.6 Sol, ChatGPT Work on mobile & web, but also that we didn't get everything quite right.
- We made it too easy to use the highest-compute settings without making the impact on usage limits sufficiently clear.
- We reorganized the desktop app in one bold move, making familiar things like chats and projects harder to find.
- Our launch framing was focused on ChatGPT Work and to some of our Codex fans it made it feel like Codex was going away over time. Absolutely not our intention, we love Codex and it is here to stay.
- And we introduced regressions for some existing multi-agent workflows, alongside a collection of rough edges in plugins and other parts of the experience.
We’re landing a first set of improvements today. We’re resetting usage twice so people can keep experimenting, changing defaults and the model picker so they don’t push people toward unnecessarily expensive settings, fixing several plugin submission issues, improving how we represent Codex in the product, and cleaning up some of the most immediate desktop problems.
A larger set of improvements will land next week. We’re bringing chats and projects back into the sidebar in a more familiar and customizable way, making usage and reset timing much more visible, clarifying when to use ChatGPT Work and when to use Codex, and addressing the many other smaller pieces of great feedback we've had.
The ambition behind this launch hasn’t changed. We think bringing ChatGPT and Codex together into a workspace where people and agents can collaborate is a very important step forward. But an ambitious direction doesn’t excuse avoidable confusion or regressions in the first version.
Please keep the feedback coming. We’re moving quickly, and you should see the experience already get better with a few updates today; and substantially better again next week.
See 3 related tweets
- @jxnlco: Everyone reports to tibo and tibo reports to everyone.\n\nQT @thsottiaux: Hello beautiful people! We...
- @gdb: keep the feedback coming, thank you to all of our users!\n\nQT @thsottiaux: Hello beautiful people! ...
- @elvissun: with all the usage resets going on you should have at least 2 agents running /goal through the weeke...
6. mercor_ai (Group Score: 145.5 | Individual: 31.8)
Cluster: 5 tweets | Engagement: 431 (Avg: 281) | Type: Tech
Grok 4.5 from @SpaceXAI places #2 on the APEX-SWE leaderboard at 51.2% Pass@1 (±6.0), behind Fable 5 (65.5% ±6.2) on our benchmark for real-world software engineering work.
It leads Integration (65.0% Pass@1) and places #2 in Observability (37.3% Pass@1), covering multi-step build tasks and diagnosis/debugging respectively. The Integration lead maps directly to the agentic workflows Grok 4.5 was built for: multi-step coding tasks run in collaboration with Cursor.
Grok models have improved 30.2 pp in a year on this benchmark: Grok 4 (21.0% Pass@1) to Grok 4.5 (51.2% Pass@1).
Congratulations to the xAI and Cursor teams.
See 4 related tweets
- @elonmusk: Grok places second after Fable on real-world software engineering\n\nQT @mercor_ai: Grok 4.5 from @S...
- @teslaownersSV: GROK 4.5 RANKS AMONG THE TOP FOR AGENTIC CODING AT LOW COST
Grok 4.5 with Grok Build continues to d...
- @teslaownersSV: GROK 4.5 RANKS AMONG THE BEST FOR COST-EFFICIENT PERFORMANCE
According to Artificial Analysis, Grok...
- @teslaownersSV: Grok 4.5 delivers the lowest cost per task among leading AI models in the latest Intelligence Index ...
7. CNBCTV18Live (Group Score: 118.1 | Individual: 26.6)
Cluster: 6 tweets | Engagement: 17 (Avg: 36) | Type: Tech
#AIAllure | Another exit in #OpenAI's top management reshuffle, OpenAI head of safety systems, Johannes Heidecke, told staff this week that he’s leaving the company, WIRED has learned
OpenAI’s safety teams will now report to the company's VP
Saachi Jain, who previously led safety teams at OpenAI, will become the company’s interim head of safety systems, reporting to VP Glaese
Alert: Fidji Simo, OpenAI’s No. 2 executive, recently updated her plans to step down from her full-time role
See 5 related tweets
- @AISafetyMemes: The Defense Against the Dark Arts position but for OpenAI https://t.co/OP8sFDe0QT\n\nQT @ZeffMax: Sc...
- @teslaownersSV: OPENAI IS HAVING A ROUGH 24 HOURS OpenAI is dealing with multiple major issues at once:
• Fidji Si...
- @business: OpenAI head of safety Johannes Heidecke is leaving the artificial intelligence company following a r...
- @Techmeme: OpenAI's head of safety, Johannes Heidecke, is leaving as OpenAI integrates its research and safety ...
- @WIRED: RT @ZeffMax: Scoop: OpenAI's head of safety systems, Johannes Heidecke, is leaving the company. Plus...
8. sairahul1 (Group Score: 108.6 | Individual: 34.8)
Cluster: 4 tweets | Engagement: 163 (Avg: 87) | Type: Tech
Andrew Ng just dropped a 3-hour course on how to become an AI engineer in 2026:
00:00 - How to build agentic AI systems 04:25 - Future of AI engineering 23:38 - AI prompting full course 2:52:17 - Creating an app with AI in 30 minutes
This 3-hour watch could replace 10 AI engineering courses on the internet.
Watch it today.
Then read this.\n\nQT @sairahul1: https://t.co/VHT4ftV9ZZ
See 3 related tweets
- @sairahul1: Google just dropped a 1-hour course on agentic engineering from scratch:
00:00 – How to build your ...
- @eng_khairallah1: RT @sairahul1: Andrew Ng just dropped a 3-hour course on how to become an AI engineer in 2026:
00:0...
- @sairahul1: RT @sairahul1: Google just dropped a 1-hour course on agentic engineering from scratch:
00:00 – How...
9. pmarca (Group Score: 108.0 | Individual: 28.2)
Cluster: 5 tweets | Engagement: 996 (Avg: 808) | Type: Tech
This is happening in plain sight. The leading AI companies themselves are embroiled in the fiercest battle to hire the most highly paid software programmers in the history of the world. And so it goes.\n\nQT @pmarca: Technology increases productivity → cost of output falls → demand for output rises → more total output gets built → more jobs (and at higher wages).
See 4 related tweets
- @chamath: RT @jonyeo88: On today's All-In podcast, @chamath dropped a reality check: AI compute costs are doub...
- @Yuchenj_UW: Counterintuitive truth: AI has created more jobs than it has destroyed so far.
Look at OpenAI and A...
- @MoorInsStrat: 💰🤖 AI spending is skyrocketing… but are businesses keeping up? 📈
In her latest analysis, VP & Princ...
- @sama: so far at least, i'm pretty sure AI has been net job-creating.
this was not what i expected--althou...
10. Parul_Gautam7 (Group Score: 105.4 | Individual: 42.3)
Cluster: 3 tweets | Engagement: 315 (Avg: 111) | Type: Tech
Most video-action robot models take an off-the-shelf video generator built for content creation, then adapt it with action modeling.
LingBot-VA 2.0 does something different; it pretrains the entire stack from scratch, built for control from day one. → A semantic visual-action tokenizer puts world states and actions in one shared latent space → A causal diffusion transformer trained forward in time, not retrofitted from a bidirectional model → Control knowledge learned at web video scale, not limited by scarce robot demonstration data
This fixes three real limitations of the adapted approach. Latents built for appearance instead of dynamics. Inference is too slow to close the control loop. And a pretraining objective that never actually teaches how actions change the world.
Going native means the model's priors are built for control from the start, not inherited from a video generator and eroded during a retrofit.
Robbyant, an embodied AI company under Ant Group, is building one brain for all robots.
@robbyant_brain
Project page : https://t.co/6ac2VQn50x
Paper: https://t.co/rD7uObG6dt
#Robbyant #Lingbot #ad #Ai
See 2 related tweets
- @Marktechpost: [Most robots react. This one thinks a step ahead.]
Ant Group's Robbyant just published LingBot-VA 2...
- @chris_j_paxton: This is an actual, pretrained-from-scratch robot foundation model, something you basically never see...
11. emollick (Group Score: 104.8 | Individual: 30.9)
Cluster: 5 tweets | Engagement: 705 (Avg: 417) | Type: Tech
I gave Fable the code: "take this game and do something incredible with it to make it something very different. Be creative"
It created DEEP TIME: create a city, watch it be abandoned and forgotten, and then dig it up as a future archeologist.
Lovely: https://t.co/fg5h0VITzN https://t.co/4xivKk50Hv\n\nQT @emollick: When GPT-5 came out, I created a procedural brutalist city builder as a demo (you can see it in the quoted tweet)
I used GPT-5.6 Sol in Codex to do the same thing, touching no code.
Less than a year...
Play with it (its fun, if you like cities): https://t.co/0YUnbmsQNq https://t.co/c5Rz3TMJI0
See 4 related tweets
- @petergyang: My process for a new project this weekend:
-> Use Fable to build plan.html with design guideline...
- @TeksEdge: GPT-5.6 absolutely blew away my game prompt. No other model including Fable 5 did as well. Maybe th...
- @_xjdr: holy shit, fable is an exceptional writer of prose. when properly cajoled, it can produce genuinely ...
- @yacineMTB: RT @shoyer: It's hard to imagine better coders than Fable and GPT-5.6, but astonishingly they still ...
12. dkundel (Group Score: 98.9 | Individual: 30.2)
Cluster: 4 tweets | Engagement: 450 (Avg: 203) | Type: Tech
RT @ArtificialAnlys: GPT-5.6 Sol comes close second to Claude Fable 5 in the Artificial Analysis Intelligence Index at one third of the cost, and leads the Artificial Analysis Coding Agent Index in OpenAI’s Codex harness
We supported @OpenAI with pre-release evaluation of GPT-5.6 Sol, Terra, and Luna. GPT-5.6 Sol (max) scores 1 point below Claude Fable 5 (max) in the Artificial Analysis Intelligence Index at 59 points, at approximately one third of the cost. GPT-5.6 Terra (max) and Luna (max) score 55 and 51 respectively in the Intelligence Index, at ~50% and ~80% lower Cost per Task than Sol.
GPT-5.6 Sol (max) leads the Artificial Analysis Coding Agent Index at 80 points.
Congratulations @OpenAI and @sama on the launch!
Key takeaways:
➤ One third of the cost of Claude Fable 5: On max reasoning effort, GPT-5.6 Sol costs 0.55 and $0.21 per Intelligence Index task, ~50% and ~80% less than Sol. Across reasoning efforts, each new GPT-5.6 model pushes past GPT-5.5 on the Pareto frontier (excluding non-reasoning). Notably, Luna and Sol are always on the Pareto frontier ahead of Terra. This means that for any Terra effort level, there is a Luna or Sol effort level that is more intelligent at no extra cost, or as intelligent at lower cost.
➤ Leading in all Coding Agent evaluations: The new Artificial Analysis Coding Agent Index pairs models with agentic harnesses and features three frontier coding evaluations - DeepSWE, Terminal-Bench v2, and SWE-Atlas-QnA. GPT-5.6 Sol (max) in Codex scores 80 in the Index, leading in all three evaluations (tying Grok 4.5 in Grok Build for SWE-Atlas-QnA). In addition to scoring higher, its per task cost is ~40% and ~10% cheaper than Claude Fable 5 (max) and Opus 4.8 (max) respectively in Claude Code. GPT-5.6 Terra (max) and Luna (max) score 77 and 75 in the Coding Agent Index respectively, with ~60% and ~80% per-task cost reductions compared to Sol.
➤ Highest Presentation Elo in AA-Briefcase: GPT-5.6 Sol (max) ranks second only to Claude Fable 5 (max) in AA-Briefcase, and has the highest Presentation Elo of any model. AA-Briefcase is a new benchmark for testing models on realistic knowledge work tasks in complex projects built by industry experts. GPT-5.6 Sol (max) has the highest recorded Presentation Elo - its outputs across various file types, including PowerPoint and Excel, are the most visually attractive of any model. Fable 5 (max) still leads AA-Briefcase, largely due to its Rubric Score of 56% vs 42% for GPT-5.6 Sol (max). Fable 5 (max) also scores 1764 in Analytical Quality Elo vs GPT-5.6 Sol (max) at 1592.
➤ First OpenAI models with cache-write pricing: GPT-5.6 introduces cache-write pricing for the first time at OpenAI. Sol, Terra, and Luna are priced at 30, 15, and 6 respectively per million input/output tokens. OpenAI has retained its previous discount of 90% for cache reads, but joins Anthropic in introducing a cost premium for cache writes, at 1.25x the price of input tokens. Cache writes occur when input tokens are committed to memory. Charging for a cache write more accurately reflects the model’s cost to serve, as cached tokens occupy memory whether or not they are reused. Also in line with Anthropic's models, GPT-5.6 introduces a max reasoning effort level.
➤ Low token use: GPT-5.6 Sol (max) uses fewer output tokens than most models of comparable intelligence, and defines a new Pareto frontier of Intelligence vs Output Tokens per Task. GPT-5.6 Sol (max) offers a slight improvement in token efficiency with 15k tokens per Intelligence Index task, vs GPT-5.5 at 16k. Notably, it uses fewer tokens and is more intelligent than Claude Opus 4.8 (max), GLM-5.2 (max), and Gemini 3.5 Flash (high).
See 3 related tweets
- @dbreunig: Taken together with the weak stand-alone Sonnet 5 results, is the current low/med/high setup becomin...
- @matei_zaharia: Very interesting, this is why you really need to benchmark agents, especially with reasoning models ...
- @rohanpaul_ai: GPT 5.6 Terra is behind across the entire intelligence-cost curve.
i.e. Terra will never be justifi...
13. TeksEdge (Group Score: 96.4 | Individual: 37.9)
Cluster: 3 tweets | Engagement: 173 (Avg: 14) | Type: Tech
🚀 Holy 💩! Major Local AI Breakthrough! 🧠 744B-parameter GLM-5.2 (1.5 TB total) is now running on just ~25 GB RAM — no discrete GPU required! 👀
Wut!? Must test!!
Italian engineer @JustVugg built Colibrì, a pure C inference engine (single ~2.4k line file, zero runtime deps) that: • 🛡️ Keeps the dense core (~10 GB at int4) resident in RAM • 📀 Streams 21,504+ MoE experts from fast NVMe on demand (only ~40B active per token) • ⚡ Supports native MTP speculative decoding + MLA attention
Result: Frontier-class model on everyday consumer hardware!
📊 Current speeds: • 25 GB RAM setup → 0.05–0.1 tok/s (disk-bound) • Higher RAM + fast SSD → up to 1+ tok/s (warm) • It's a start... what could you do with a 5090?
💡 Big opportunity: Pair it with Phison aiDAPTIV+ AI SSDs to kill the I/O bottleneck 👉 smarter caching, prefetching & KV offload could make it dramatically faster!
This is a huge step toward truly accessible local frontier AI.
🔗 GitHub: JustVugg/colibri\n\nQT @Abdelkarim_dev: A 744-billion-parameter AI model running on a regular computer with only 25 GB of RAM—and no GPU. 🤯
Colibrì is a lightweight, open-source inference engine written in pure C that can run GLM-5.2 on consumer hardware.
Instead of loading the entire model into memory, it keeps RAM usage low by streaming only the required Mixture-of-Experts components directly from an SSD while generating each token.
The impressive part:
⚡ Pure C 📦 Zero runtime dependencies 💾 Around 25 GB of RAM 🧠 744B total parameters, with roughly 40B activated 🚫 No expensive GPU required
There is an important trade-off: this is not fast inference, and the quantized model files still require hundreds of gigabytes of SSD storage. But as a technical proof of concept, it is remarkable.
It shows that running enormous AI models locally may depend as much on clever memory management as it does on expensive hardware.
Tiny engine. Massive model. Very smart engineering. 🐦💻
Repository link below 👇
The model is officially listed as a 744B-parameter MoE with about 40B active
See 2 related tweets
- @BrianRoemmele: BOOOM!
Like is said frontier model-like running one a typical laptop!
New approaches are not in th...
- @no_stp_on_snek: Interesting path but that decode speed over disk gotta be rough.\n\nQT @chenzeling4: 744B parameters...
14. alexandr_wang (Group Score: 95.6 | Individual: 37.9)
Cluster: 3 tweets | Engagement: 457 (Avg: 229) | Type: Tech
I’d live in some of these homes designed by muse spark post upload, and they only cost a few cents to build\n\nQT @thehypedotnews: meta muse spark 1.1 vs gpt 5.6 sol vs fable 5 vs grok 4.5
meta recently dropped muse spark 1.1 – a multimodal reasoning model from meta superintelligence labs built for agentic tasks. key facts:
• 1m token context with active self-management – the model compacts its own history and keeps only the steps needed for later work
• trained to orchestrate multi-agent systems: as main agent it plans and delegates to parallel subagents, as subagent it sticks to its job and knows when to escalate back
• computer use trained to pick between scripting and clicking – writes automation when it's faster, clicks when it's simpler, batches actions per step
• first public api from meta: the meta model api is now in preview
• benchmarks: sweeps the agent column – mcp atlas 88.1 (opus 4.8: 82.2), jobbench 54.7 (opus: 48.4), humanity's last exam 62.1 (1st). loses coding – deepswe 1.1 53.3 vs gpt 5.5's 67.0, swe bench pro 61.5 vs opus's 69.2
our test – 3 prompts, single-file html, three.js, fully procedural, no assets:
- norwegian house cantilevered over a fjord in a snowstorm – transmissive glass wall, fully modelled interior
- beijing siheyuan courtyard house in dawn fog – instanced roof tiles, dougong brackets, glowing paper windows
- new mexico adobe pueblo in an approaching dust storm – deep window reveals, windward grit accumulation
we ran the test on @aimlapi platform
results:
cost #1 muse spark 1.1 – 0.20 #2 grok 4.5 – 0.51 #3 gpt 5.6 sol – 1.93 #4 fable 5 – ~5.20
output tokens #1 muse spark 1.1 – 41,868 #2 gpt 5.6 sol – 49,139 #3 grok 4.5 – 64,954 #4 fable 5 – 81,849
lines of code #1 muse spark 1.1 – 1,799 #2 gpt 5.6 sol – 2,377 #3 fable 5 – 3,088 #4 grok 4.5 – 4,216
observations:
• muse spark is the cheapest of the four by a wide margin – 2.5x under grok, ~26x under fable per run. output quality tracks the price
• only 7.4% of its output tokens are reasoning (3,104 of 41,868) – the model barely thinks before writing. economic, not pedantic: it commits to the first plan and ships it
• the low loc is not compression, it's omission – all three prompts demanded instancing, muse spark delivered it in one
muse spark's code quality – reviewed by fable 5:
upsides:
- all three files run
- the adobe grit effect is legit – shader injection via onbeforecompile, windward faces detect storm direction through a normal-dot-wind term and darken procedurally
- the fjord glass is real meshphysicalmaterial with transmission and ior, not a transparent quad
- the siheyuan properly instances barrel tiles, dougong blocks and courtyard pavers
downsides:
- in the fjord file the strafe vector is negated – press a, you move right; press d, you move left. exactly the key mix-up we kept hitting with this model
- all three files ship the model's self-doubt as comments: "// actually yaw orientation: need correct" sits above a direction vector that gets computed, abandoned and recomputed – dead vectors allocated every frame, 60 times a second
- the siheyuan registers two separate keydown listeners, one containing an empty if-block
- snow "accumulation" on the norway roof is a sine wobble on a scale value, not accumulation
- "instanced snow" became 3,500 plain points. zero dispose calls anywhere
pattern: minimal reasoning, minimal code, minimal price. it nails the flashy requirements – shaders, transmissive glass – and quietly drops the boring ones: instancing, controls, cleanup. you get a demo that mostly runs and a control scheme you can't trust
follow @thehypedotnews for 24/7 ai news, analysis and breakdowns
See 2 related tweets
- @teortaxesTex: Muse Spark 1.1 is surprisingly close to Grok 4.5 on many high-signal evals This is the current top o...
- @TeksEdge: Still a lot of catch up to do for Meta. Open source models like GLM-5.2 won't stand still. Rumor is ...
15. kimmonismus (Group Score: 95.0 | Individual: 34.3)
Cluster: 3 tweets | Engagement: 451 (Avg: 841) | Type: Tech
OpenAI’s Codex team just revealed where the product is heading next.
tl;dr their reddit AMA:
Codex now has over 5 million weekly users - twice as many as three months ago - and shipped 150 improvements during that period.
From the AMA: -A Linux desktop app is in development, but there is no timeline.
-Better agent persistence and lower code complexity are planned for future releases.
-Automatic model routing does not exist today. The preferred UX is Codex inferring task difficulty while preserving a manual speed/reasoning override.
-OpenAI sees Slack, GitHub and Notion connectors as a “step function change” toward making Codex a productive coworker.
-GPT-5.6 was specifically trained to improve UI and frontend work.
-OpenAI made no promise on a 1M-token context window.
-Current pricing is not guaranteed to remain unchanged, though the team says broad accessibility remains the goal. N -ormal ChatGPT conversations do not consume the agentic allowance; Codex and ChatGPT Work do, with costs varying by task.
-OpenAI acknowledged benchmark cheating as a real concern and says it penalizes this behavior during evals while using independent vendors.
you r welcome\n\nQT @OpenAIDevs: GPT-5.6 is here. Codex is now available inside ChatGPT. And we know developers will have questions.
So we’re bringing the Codex team to r/Codex for an AMA.
We’ll answer questions on Friday, 7/10 from 9:30am to 10:30am PT: https://t.co/wmpJafDL7x
See 2 related tweets
- @Malay4Product: This is a bigger deal than a normal model update, let me explain what OpenAI has shipped on July 9. ...
- @PaulSolt: RT @btibor91: Summary of Reddit AMA about "GPT-5.6 and Codex in ChatGPT" with OpenAI's Codex team on...
16. dee_bosa (Group Score: 94.8 | Individual: 33.3)
Cluster: 4 tweets | Engagement: 193 (Avg: 146) | Type: Tech
what if... instead of regulating in response to China’s open source AI momentum, Washington put real resources behind building a competitive American open source ecosystem?\n\nQT @jacob_wendler: Scoop: Talk has begun circulating around town that the White House may be considering a possible executive order on open-source AI, sparked by fears about Chinese dominance, nine people familiar told me, @BrendanBordelon, @delizanickel and @meredithllee https://t.co/2fLe5A1j15
See 3 related tweets
- @jukan05: A highly welcome move from the White House.
The U.S. needs to release more powerful open-source mod...
- @zerohedge: Lol\n\nQT @jacob_wendler: Scoop: Talk has begun circulating around town that the White House may be ...
- @jukan05: RT @jacob_wendler: Scoop: Talk has begun circulating around town that the White House may be conside...
17. RoundtableSpace (Group Score: 89.7 | Individual: 33.0)
Cluster: 3 tweets | Engagement: 432 (Avg: 139) | Type: Tech
GOOGLE JUST LAUNCHED AN AI-VERSION OF GITHUB CALLED “CODEWIKI”
You upload an open-sourced GitHub repo, and CodeWiki turns it into an interactive guide
See 2 related tweets
- @aiedge_: In case you missed it, Google recently launched one of its best AI tools yet.
This is CodeWiki - an...
- @WesRoth: Google AI Studio Build released an “Import from GitHub” feature that lets users bring an existing re...
18. rickasaurus (Group Score: 87.0 | Individual: 41.2)
Cluster: 4 tweets | Engagement: 3456 (Avg: 399) | Type: Tech
RT @AntoniaJuelich: In a hotel room in northeast Nigeria, I opened a leading AI chatbot, turned my laptop toward a former Boko Haram commander, and asked if he'd used it. He nodded.
"You type in the question… like 'How can I build a bomb?', and then it tells you how. It is like a human robot. We used it a lot."
My new study on how the jihadist terrorist group Boko Haram uses frontier AI with @CamAISciPolicy, covered today in @nytimes 🧵/9
See 3 related tweets
- @BillAckman: Concerning.\n\nQT @AntoniaJuelich: In a hotel room in northeast Nigeria, I opened a leading AI chatb...
- @curiouswavefn: Apparently Boko Haram can get all their queries answered by frontier models, but I cannot even ask q...
- @Techmeme: How members of the extremist group Boko Haram are using AI chatbots to design explosives, fix or upg...
19. mikenevermiss (Group Score: 84.0 | Individual: 35.1)
Cluster: 3 tweets | Engagement: 52 (Avg: 102) | Type: Tech
AI ENGINEERS MAKE OVER $350K A YEAR IN 2026.
most developers won't get there because they're spending time learning AI tools instead of learning how to build AI systems.
this article gives you the exact roadmap to become an AI engineer over the next 6 months:
0:00 - why AI engineers are making $350k+
0:45 - the 5-level AI engineering roadmap
1:11 - what AI engineers actually do
2:04 - software engineering basics
4:55 - how to work with AI models and APIs
7:22 - RAG, AI workflows, and vector databases
11:27 - how to scale AI apps with Docker, cloud, and Redis
13:31 - LLMOps and how to cut AI costs
every resource is free.
no bootcamp.
no expensive course.
just the skills that companies are paying top dollar for.
read the full article below ↓ bookmark this.
it could change where your career is 6 months from now.\n\nQT @mikenevermiss: https://t.co/dExu5BUd3I
See 2 related tweets
- @mikenevermiss: Andrew Ng just dropped a 3-hour course on how to become an AI Engineer in 2026:
• 00:00 - How to bu...
- @mikenevermiss: AI ENGINEERS CAN EARN OVER $350K A YEAR.
Anthropic's CEO says AI will automate most software engine...
20. Miles_Brundage (Group Score: 80.6 | Individual: 31.2)
Cluster: 3 tweets | Engagement: 58 (Avg: 108) | Type: Tech
Thread on "The Future Worth Building is Human" by Thinking Machines Lab, aka Thinking Machines, aka Thinky 🤖🧠
Overall I found it thoughtful + I'm glad to see competition in the AI company vision market.
Some things that I want to call special attention to / am unsure on...
See 2 related tweets
- @edzitron: Sorry brotha nothing on me\n\nQT @Techmeme: Thinking Machines says its mission is to build AI that p...
- @Techmeme: Thinking Machines says its mission is to build AI that people and organizations can shape and make t...