Image created with gemini-2.5-flash-image with claude-sonnet-4-5. Image prompt: Photorealistic Times Square at dusk dominated by massive glowing red billboards displaying cybersecurity warnings, padlock icons, firewall diagrams, and threat alerts, pedestrians below bathed in red-amber light, digital glitch effects cascading down building facades, ominous and vigilant atmosphere, shot from street level looking up at the towering security-themed displays.

We’re at an inflection point in AI’s impact on cybersecurity. Claude now outperforms human teams in some cybersecurity competitions, and helps teams discover and fix code vulnerabilities. At the same time, attackers are using AI to expand their operations. https://x.com/AnthropicAI/status/1974199155657748868

AI for Cyber Defenders \ red.anthropic.com https://red.anthropic.com/2025/ai-for-cyber-defenders/

Very excited to see the Tinker release! @pcmoritz and I had a chance to experiment with the API. It does a nice job of providing flexibility while abstracting away GPU handling. Here’s a simple example showing how to generate synthetic data and fine tune a text to SQL model.”” / X https://x.com/robertnishihara/status/1973455582603649430

Tinker provides an abstraction layer that is the right one for post-training R&D — it’s the infrastructure I’ve always wanted. I’m excited to see what people build with it. “”Civilization advances by extending the number of important operations which we can perform without”” / X https://x.com/johnschulman2/status/1973450054238347314

A flexible API for fine-tuning LMs – Tinker by @thinkymachines Write a simple CPU-only script, and it runs your exact training loop on distributed GPUs. You can fine-tune open models like Llama and Qwen, up to large MoE (Qwen3-235B-A22B), switching them by changing only one https://x.com/TheTuringPost/status/1973827605448306883

Really excited and proud to see Qwen models are in the first batch of supported models for the tinker service! 🤩 we will continue to release great models to grow research in the community 😎 https://x.com/wzhao_nlp/status/1973603599616974970

I’ve been using Tinker at Redwood Research to RL-train long-context models like Qwen3-32B on difficult AI control tasks – specifically teaching models to write unsuspicious backdoors in code similar to the AI control paper. Early stages but seeing some interesting backdoors 👀”” / X https://x.com/ejcgan/status/1973449963259699284

It turns out that the AI jagged frontier worked as a reverse salient, a term from the history of science for a technology or process that holds back the whole system & thus a focus of development. Math & planning were reverse salients, so they have seen the most improvement. https://x.com/emollick/status/1973148208894451908

I had the chance to try @thinkymachines’ Tinker API for the past couple weeks. Some early impressions: Very hackable & lifts a lot of the LLM training burden, a great fit for researchers who want to focus on algs + data, not infra. My research is in RL, and many RL fine-tuning”” / X https://x.com/tyler_griggs_/status/1973450947218252224

Tinker is cool. If you’re a researcher/developer, tinker dramatically simplifies LLM post-training. You retain 90% of algorithmic creative control (usually related to data, loss function, the algorithm) while tinker handles the hard parts that you usually want to touch much less”” / X https://x.com/karpathy/status/1973468610917179630

🚀With early access to Tinker, we matched full-parameter SFT performance as in Goedel-Prover V2 (32B) (on the same 20% data) using LoRA + 20% of the data. 📊MiniF2F Pass@32 ≈ 81 (20% SFT). Next: full-scale training + RL. This is something that previously took a lot more effort”” / X https://x.com/chijinML/status/1973451597393883451

thinking-machines-lab/tinker-cookbook: Post-training with Tinker https://github.com/thinking-machines-lab/tinker-cookbook

[1 Oct 2025] Thinking Machines’ Tinker: LoRA based LLM fine-tuning API https://x.com/Smol_AI/status/1973622595124863044

Announcing Tinker – Thinking Machines Lab https://thinkingmachines.ai/blog/announcing-tinker/

One interesting “”fundamental”” reason for Tinker today is the rise of MoE. Whereas hackers used to deploy llama3-70B efficiently on one node, modern deployments of MoE models require large multinode deployments for efficiency. The underlying reason? Arithmetic intensity. (1/5) https://x.com/cHHillee/status/1973469947889422539

Very excited to see the Tinker release by @thinkymachines! @robertnishihara and I had a chance to experiment with the API, see https://x.com/pcmoritz/status/1973456462346424641

Tinker – Thinking Machines Lab https://thinkingmachines.ai/tinker/

OpenAI’s social creation & consumption experiment is here. Unique handling of identity – if you opt-in others can use your likeness in their creations. You’re notified even if it’s used in a draft post & “liveness checks” are done to prevent impersonation. https://x.com/bilawalsidhu/status/1973103500511871277

Had fun being in Germany to launch a sovereign cloud offering with SAP and Microsoft; important to us to help governments use our frontier models.”” / X https://x.com/sama/status/1971433413086499044

Unitree CEO Wang Xingxing at a Trade Fair in Hangzhou on Saturday: ⦿ Unitree R1 will become the world’s best-selling humanoid robot next year. ⦿ In the first half of this year, the domestic robot industry grew an average rate of 50% to 100% for Chinese intelligent https://x.com/TheHumanoidHub/status/1973158573317501243

Unitree CEO Wang Xingxing expects R1 to be the world’s best-selling humanoid robot next year. Won’t shock anyone if it happens. The company announced the starting price of $5,900 but even at $12k this will sell like hot cakes https://x.com/TheHumanoidHub/status/1973452915366044096

Governor Newsom signs SB 53, advancing California’s world-leading artificial intelligence industry | Governor of California https://www.gov.ca.gov/2025/09/29/governor-newsom-signs-sb-53-advancing-californias-world-leading-artificial-intelligence-industry/

Anyone who sees this video can instantly grasp the (at least) potential for malicious use. And yet nobody with any power (either in the public or at the corporate level) has anything to say (let alone do) to address it, or even acknowledge it.”” / X https://x.com/TheStalwart/status/1973372434133950665

I get data sovereignty in some cases, but there is just no way for new countries to join the frontier model race as long as scaling (in any sense) matters. There is no sovereign model. You will be dependent on the production of Chinese (or US or French) open models as a base.”” / X https://x.com/emollick/status/1972018517919826099

We’ve focused on improving Claude’s skills in defensive cybersecurity. The results of this are visible in Claude Sonnet 4.5, which is comparable or superior to Opus 4.1 in cybersecurity tasks—yet both faster and cheaper. Read more: https://x.com/AnthropicAI/status/1974199158929305738

Anthropic to triple international workforce in global AI push https://www.cnbc.com/2025/09/26/anthropic-global-ai-hiring-spree.html

Daiwa Securities is hiring startup Sakana AI to build an AI tool analyzing investor profiles, joining other firms adopting the technology (Bloomberg: https://x.com/SakanaAILabs/status/1974109165623853365

We are pleased to announce our partnership with Daiwa Securities, a major financial services firm in Japan. https://x.com/SakanaAILabs/status/1973935631354245286

A senior government official of the UAE, Abdulla M. Alhamed, met Optimus and Elon at Tesla HQ in California. https://x.com/TheHumanoidHub/status/1972093872177401983

Security researchers have discovered that Unitree robots, like the humanoids G1 and H1 and quadrupeds Go2 and B2, suffer from a severe security hole in their Bluetooth Low Energy (BLE) setup for the Wi-Fi configuration interface. This flaw lets attackers perform command https://x.com/TheHumanoidHub/status/1971314995779748126

Endpoint Security for AI eBook https://www.delltechnologies.com/asset/en-us/solutions/business-solutions/briefs-summaries/endpoint-security-for-ai-ebook.pdf

Own AI Securely with SANS | SANS Institute https://www.sans.org/mlp/ai-security-blueprint

AI Security Starts Here | SANS Institute https://www.sans.org/mlp/artificial-intelligence

DevSecCon: Securing the Shift to AI Native | Register for Free | Oct ’25 | Snyk https://snyk.io/events/devseccon/

Last week we found an issue with SWE-Bench, allowing agents to cheat by looking at future commits. Instead of celebrating the SWE-Bench Devs for quickly fixing the issue and being transparent, the HN crowd is dunking on them and drawing wildly inaccurate conclusions about”” / X https://x.com/TacoCohen/status/1966421688846778561

Its kind of funny that AI can definitely do most common CAPTCHAs better than humans and the reason that CAPTCHAs still work is because the big LLMs often refuse to do them. https://x.com/emollick/status/1972177860086612283

Trending

Discover more from Ethan B. Holland

Subscribe now to keep reading and get access to the full archive.

Continue reading