About This Week’s Covers
This week’s newsletter category covers are a nod to stereotypical ’80s TV infomercials.
The main cover was created using a combination of GPT-Image-1 and MidJourney. I layered them together using Photoshop.
For the rest of the covers, GPT o3 has a saved rubric and can automatically adapt to any theme I give it. I can provide a one sentence theme, and o3 automatically generates 46 cover images using the API with no supervision. All of the ideas and compositions come from GPT on its own. My prompt this week was “an over the top stereotype of an ’80s or ’90s TV infomercial screen grab.” Everything else was automated. It’s not an attempt to be amazing quality, but instead see how creative GPT can be without any help.
A few turned out pretty well! I’ve included my favorite six of the covers below:

This Week By The Numbers
Total Organized Headlines: 646
- AGI: 44 stories
- Accounting and Finance: 26 stories
- Agents and Copilots: 253 stories
- Alibaba: 7 stories
- Amazon: 9 stories
- Anthropic: 58 stories
- Apple: 19 stories
- Audio: 22 stories
- Augmented Reality (AR/VR): 36 stories
- Autonomous Vehicles: 25 stories
- Benchmarks: 44 stories
- Business and Enterprise: 133 stories
- ByteDance: 4 stories
- Chips and Hardware: 36 stories
- DeepSeek: 29 stories
- Education: 47 stories
- Ethics/Legal/Security: 146 stories
- Figure: 3 stories
- Google: 62 stories
- HuggingFace: 17 stories
- Images: 32 stories
- International: 68 stories
- Llama: 8 stories
- Locally Run: 13 stories
- Meta: 37 stories
- Microsoft: 27 stories
- Mistral: 8 stories
- Mobile: 8 stories
- Multimodal: 41 stories
- NVIDIA: 14 stories
- Open Source: 82 stories
- OpenAI: 79 stories
- Perplexity: 5 stories
- Podcasts/YouTube: 48 stories
- Publishing: 84 stories
- Qwen: 5 stories
- RAG: 5 stories
- Robotics Embodiment: 77 stories
- Safe Superintelligence: 1 story
- Science and Medicine: 43 stories
- Technical and Dev: 107 stories
- Video: 48 stories
- X: 11 stories
This Week’s Executive Summaries
I’m a few weeks behind because family and friends continue to come first. My spare time on the weekends has been spent celebrating the life of my great friend, Mike Bernstein, and rooting for my daughters in their national dance competitions.
AI has not slowed down, and the week ending June 20th has over 30 executive summaries.
The big stories are robotics and job disruption.
Nvidia CEO Jensen Huang predicts that humanoid robots will be a $50 trillion market with over 1 billion robots. Jensen is not some science fiction writer. He is not a hyperbolic venture capitalist. This is the CEO of the largest chipmaker on earth.
Andy Jassy, CEO of Amazon, announced that they expect to significantly reduce their workforce in the coming years.
OpenAI’s Greg Brockman, one of the world’s leading AI experts, expects that AI will quickly evolve from becoming our coworkers to our managers.
British Telecom told the Financial Times that AI will lead to workforce reductions.
AI coding assistants are so good that they are now comparable to hiring junior software engineers and are starting to replace salaries rather than being considered software subscriptions.
A recent study valued AI-assisted coding in the United States alone at over $10 billion annually.
OpenAI’s Codex merged over 352,000 pull requests with an 85% success rate in under 35 days.
Apple came out with artificial intelligence models that were specifically designed to be run locally on phones. More importantly, Apple launched an API that allows developers to integrate their own code into iPhone apps. This is a huge harbinger of something that I’ve been shouting from the rooftops for the past 10 months. I am convinced Apple has a plan and is not as far behind with artificial intelligence as people think.
As if Google was reading Apple’s mind, Google launched voice features that let users talk with search results instead of just reading them. This feels like a direct assault against both Apple Siri and Amazon Echo as interfaces become more conversational. Google’s voice integration can run in the background across apps and link to web sources to be clicked on later. The fact that it subtly works across apps without having to open them is another huge development that is pushing towards what I saw from Apple’s Ferret model a year ago.
Microsoft released an upgrade to its assistant that can see and analyze what’s on your computer screen to give real-time guidance or directions to troubleshoot or modify settings and use software. For example, someone could have Adobe Lightroom open and ask the assistant how to sharpen or brighten a photo, and the assistant will be able to see the screen and walk the user through how to make the desired results. Clearly, this is one step away from the assistant simply doing the work itself. OpenAI’s screen sharing mode has been able to do this for a few months, and it’s worth trying if you have the app.
A lot of people push back on my assertion that voice will be the next big interface. However, I feel voice has incredible opportunities to move people off of keyboards. I understand that business travelers are already obnoxious, but for home use, Airpods, and driving, voice is something that I would embrace. It’s easier to vocalize thoughts than to type them.
Continuing the voice trend, XAI is testing a voice conversation tool within Grok. X is also launching a tasks feature similar to OpenAI’s scheduler.
Going back to the robotic theme of the week, Nvidia launched new versions of their physics simulation environments where developers can build, train, and test robots before deploying them in the real world.
Figure Robotics shared a video of their humanoid robot performing open-ended logistics work on an assembly line for over 60 minutes with 95% accuracy. The average time for the robot to process a package has gone down from 6 to 4 seconds.
Meta unveiled a new learning model for robots that can understand and predict physical interactions by watching videos and learning, similarly to how children learn through observation.
Quite a few defense and national security contracts popped up last week.
OpenAI secured a $200 million Department of Defense contract to provide AI systems for national security and combat. Anduril is partnering with German defense companies to manufacture military drones for European markets.
The biggest defense news is that the US Army appointed executives from Palantir, Meta, and OpenAI as Lieutenant Colonels.
US officials reported that they are using AI tools more to accelerate analysis and decision-making across government operations.
The UK deployed a system called Extract that converts complex planning documents, including handwritten notes and blurry maps, into digital data within 40 seconds.
Japan’s largest bank partnered with AI company Sakana to automate banking document creation.
Russia’s Sberbank announced plans to create their own AI system with advanced reasoning capabilities.
The world of legacy user interfaces is starting to erode.
Google demonstrated a new ability by Gemini 2.5 to create custom user interfaces on the fly based on the user’s needs. In real time, the AI writes code for an entirely new interface as users click on the buttons. Everything is optimized for a fluid user experience as opposed to a predefined experience of choices and designs. There’s a fun demo video that’s worth watching.
Salesforce released a marketing suite of tools that can handle routine marketing tasks like building audience segments, writing email copy, and managing ad performance. These agents can work independently or assist marketing managers, and take directions like building a campaign or creating personalized offers.
OpenAI rolled out a record mode for ChatGPT and can capture meetings and voice notes. This is a perfect example of how entire software companies will be eaten by a new feature in a frontier model…just like physical GPS devices in cars have been replaced by an app on your phone.
Last week, Apple came out with a paper that claimed artificial intelligence models can’t reason. People are using Claude and GPT to create counterpoint scientific papers with citations and references to create an argument that humans can’t reason. These are very fun ways to troll Apple, but also a bit existential because the papers are quite compelling.
A medical model was released that can answer questions as well as human physicians on healthcare benchmarks. What’s really wild about this model is that it has 70% fewer parameters and can run on a laptop locally. We’re getting to the point where everyone will be able to have a physician in their pocket without even being connected to the cloud.
OpenAI announced that they are implementing security measures as their models approach high levels of capability in biology. These models could accelerate drug discovery and vaccine development but could also potentially assist in creating biological weapons. OpenAI is going to host a biodefense summit in July.
ByteDance, the parent company of TikTok, launched an AI system that creates high-quality videos from text prompts in less than one minute, a 10X speed improvement over their previous model.
Old-school image generation company, Midjourney released its first video model that can transform a static image into a five second animated clip.
Amazon released an open source security researcher agent that can learn and create custom tools on the fly. The agent combines memory storage, self reflection and tool creation and can complete full security assessments in less than 10 minutes.
In chips and hardware news, Amazon announced plans for an updated server CPU and a new AI training chip.
Apple announced that it will use AI to help design its own processors.
SoftBank pitched a $1 trillion artificial intelligence hub in Arizona.
Nvidia partnered with Mistral to build an AI cloud infrastructure using American-made chips.
In spill the tea news, a repository of files and reporting was released documenting alleged misconduct by Sam Altman with a treasure trove of negative feedback from former OpenAI employees.
All this and more in this week’s newsletter below. Don’t forget to get outside!
Nvidia CEO predicts billion-robot future worth $50 trillion
Nvidia CEO Jensen Huang told an audience in Paris that humanoid robotics could become “one of the largest industries ever,” predicting a future with a billion robots. The company estimates that physical AI systems represent a $50 trillion market opportunity as artificial intelligence moves from digital applications into real-world robotic forms. Nvidia’s bullish forecast reflects the tech industry’s growing focus on embodied AI, where software intelligence gets paired with physical hardware to perform tasks in homes, workplaces, and manufacturing facilities.
Jensen at GTC Paris keynote: “”Humanoid robotics is going to potentially be one of the largest industries ever.”” “”The idea that there would be a billion robots is a very sensible thing.”” https://x.com/TheHumanoidHub/status/1933246446297624607
Nvidia believes physical AI systems are a $50 trillion market opportunity – GamesBeat https://gamesbeat.com/nvidia-believes-physical-ai-systems-are-a-50-trillion-market-opportunity/
Amazon CEO says AI will reduce corporate workforce in coming years
Amazon CEO Andy Jassy announced that the company expects generative AI to reduce its total corporate workforce over the next few years as AI handles more routine tasks. Jassy said Amazon currently has over 1,000 AI services and applications in development, with plans to accelerate AI agent deployment across all business units to improve efficiency and customer experiences. Jassy emphasized that while some jobs will be eliminated, new types of roles will emerge as the company transforms how work gets done with AI assistance.
RT @AndrewCurran_: Amazon CEO Andy Jassy in a memo to employees earlier today: ‘in the next few years, we expect that this will reduce our…”” / X https://x.com/dilipkay/status/1935079746196451411
Update from Amazon CEO Andy Jassy on Generative AI https://www.aboutamazon.com/news/company-news/amazon-ceo-andy-jassy-on-generative-ai
Greg Brockman predicts we’ll have AI managers
OpenAI’s Greg Brockman expects AIs to go from AI coworkers to AI managers: “the AI gives you ideas and gives you tasks to do”
OpenAI’s Greg Brockman expects AIs to go from AI coworkers to AI managers: “”the AI gives you ideas and gives you tasks to do”” : r/singularity https://www.reddit.com/r/singularity/comments/1lfe3ih/openais_greg_brockman_expects_ais_to_go_from_ai/
BT Group (formerly British Telecom) CEO says AI will likely accelerate planned job cuts
BT Group’s CEO told the Financial Times that AI technology could lead to deeper workforce reductions beyond the company’s existing plan to cut over 40,000 jobs by 2030. The British telecom giant is already implementing a £3 billion cost-cutting program, but advancing AI capabilities may allow for even more automation of tasks currently done by humans.
BT boss Kirkby expects AI to deepen job cuts, FT reports | Reuters https://www.reuters.com/business/media-telecom/bt-ceo-eyes-deeper-job-cuts-ai-becomes-more-powerful-ft-reports-2025-06-15/
AI coding tools reach price parity with human developers
AI coding assistants are becoming so capable that businesses now compare their costs to hiring junior software engineers rather than traditional software subscriptions. A recent study valued AI-assisted coding in the US at $9.6-14.4 billion annually, with potential to reach $64-96 billion as productivity gains from controlled trials scale up. This represents a critical shift where AI tools transition from nice-to-have productivity boosters to viable alternatives to human labor. The implications for software engineering careers are significant, with industry observers noting that new developers will need exceptional skills to compete. OpenAI’s coding agent recently merged over 352,000 pull requests with an 85.5% success rate in just 35 days, demonstrating the scale and reliability these systems are achieving in real development workflows.
Heard this quote at a dinner recently: “”I used to think Claude Code was too expensive compared to other dev tools. But after Opus 4, I realized it’s actually very affordable compared to what you’d pay a junior software engineer.”” It feels important to note that we’re nearing a”” / X https://x.com/alexalbert__/status/1936109179594494381
Lots of interesting stuff in this paper, plus, as of the end of 2024: “the annual value of AI-assisted coding in the United States at $9.6−14.4 billion, rising to 64−96 billion if we assume higher estimates of productivity effects reported by randomized control trials””” / X https://x.com/emollick/status/1933206622483939491
If you are beginning a software engineering career, you MUST be at the the top 1%, else it will be very difficult. OpenAI’s SWE AI coding agent Codex merged 352K+ Pull Request with 85.5% success rate. And this number is just for previous 35 days. This repo tracks the opened https://x.com/rohanpaul_ai/status/1936041618554954142
Vibe Coding Is Coming for Engineering Jobs | WIRED https://www.wired.com/story/vibe-coding-engineering-apocalypse/
Apple updates its Foundation Models for devices and servers
Apple updated its Apple Foundation Models with versions designed for on-device and server use, focusing on image understanding and multilingual tasks. The company also launched a Foundation Models API that lets developers integrate the on-device model into apps on Apple Intelligence-enabled devices. The on-device version uses a compact 3 billion parameter design that outperformed similar-sized competitors in some tests, while the server version showed mixed results against leading models like GPT-4o. Apple’s approach emphasizes running AI directly on user devices rather than relying solely on cloud processing.
Welcome to Agentforce Marketing: Agentforce changes everything — again. It runs campaigns, personalizes every touchpoint, slashes costs. One agent for sales, service & marketing. The future of customer interaction starts now. ❤️🤖 https://x.com/Benioff/status/1933017834512056444
Google is coming for Siri and Alexa, adds voice conversations and audio summaries to search
Google launched two voice features that let users talk with search results instead of just reading them. Audio Overviews provides spoken summaries of search topics using Gemini AI, while Search Live enables back-and-forth voice conversations with Google’s search engine on mobile devices. Both features work hands-free, allowing people to multitask while getting information. The voice conversation feature runs in the background across apps and includes links to web sources for clicking through later.
New Audio Overviews experiment in Search coming to Labs (comin for ya Amazon and Siri!) https://blog.google/products/search/audio-overviews-search-labs/
Search Live with voice in AI Mode on Google Search https://blog.google/products/search/search-live-ai-mode/
xAI tests voice mode and automated tasks for Grok
xAI is testing two major features for its Grok chatbot: voice conversations and scheduled tasks. Voice mode now works on the web version, letting users speak to Grok and hear responses back, similar to the mobile app. The Tasks feature allows users to schedule regular searches or research queries, with options for daily monitoring of specific topics or news. Users can set up to 3 daily recurring tasks and 10 occasional ones, a significant reduction from previous limits of 10 daily and 30 occasional tasks.
xAI tests Voice Mode and Tasks on Grok web with Grok 3.5 (Amazon, Siri) https://www.testingcatalog.com/xai-tests-voice-mode-and-tasks-on-grok-web-with-signs-of-grok-3-5-integration/
UBS predicts 300 million humanoid robots by 2050
UBS analysts forecast that humanoid robots will grow from 2 million units in 2035 to 300 million by 2050, creating a market worth up to $1.7 trillion. The report cites aging populations and labor shortages as key drivers, with robots designed in human form offering better adaptability to existing workplaces and daily life. The bank expects robot costs to drop over 70% in the next two decades due to improved manufacturing scale and supply chains, making widespread adoption economically viable.
300 million humanoid robots are coming – and here are the companies that will benefit | Morningstar https://www.morningstar.com/news/marketwatch/20250618137/300-million-humanoid-robots-are-coming-and-here-are-the-companies-that-will-benefit
Nvidia releases Isaac Sim 5.0 and Isaac Lab 2.2 for robotics development
Nvidia launched early preview versions of Isaac Sim 5.0 and Isaac Lab 2.2 on GitHub for robotics developers. These open-source frameworks provide physics-based simulation environments where developers can build, train, and test AI robots before deploying them in the real world. The tools include extensions for synthetic data generation and pre-built robot models, which streamline the development process by allowing developers to create realistic training scenarios without needing physical hardware for every test phase.
#NVIDIAIsaac Sim 5.0 and Isaac Lab 2.2 are now available in early developer preview on Github. 🎉 These releases give #Robotics developers early access to cutting-edge tools to simulate, train, and validate robots in a physics-based simulation environment. What’s new? https://x.com/NVIDIARobotics/status/1934768379652665403
Nvidia launched Issac Sim 5.0 and Issac Lab 2.2 in early preview on GitHub These open frameworks now come with extensions for synthetic data generation and robot models — streamlining how devs build, train, and test AI robots in physics-based simulations https://x.com/rowancheung/status/1934881540263018666
Amazon tests humanoid robots for package delivery
Amazon is building an indoor obstacle course at its San Francisco office to test humanoid robots that could eventually handle delivery tasks. The company is developing AI software to power these robots while planning to use hardware from other manufacturers for initial testing. The move represents Amazon’s broader push to integrate AI across its logistics network, from warehouses to final delivery. Amazon has not commented on the timeline for deployment, but the testing facility suggests the company is moving beyond warehouse automation toward replacing human delivery workers with robotic alternatives.
Amazon prepares to test humanoid robots for deliveries, The Information reports | Reuters https://archive.md/o682X
Flashback: Two weeks ago, Meta released V-JEPA 2 world model for physical reasoning
Meta unveiled an AI system that can understand and predict physical interactions by watching videos, similar to how children learn about gravity and object behavior through observation. V-JEPA 2 can watch a video of objects moving and predict what will happen next, then use that knowledge to control robots in unfamiliar environments without additional training. The model represents a step toward AI systems that can plan actions by mentally simulating outcomes, much like humans do when navigating crowded spaces or playing sports. Meta also released three new tests to measure how well AI systems understand physics from video, addressing a gap in current evaluation methods for physical reasoning capabilities.
Our vision is for AI that uses world models to adapt in new and dynamic environments and efficiently learn new skills. We’re sharing V-JEPA 2, a new world model with state-of-the-art performance in visual understanding and prediction. V-JEPA 2 is a 1.2 billion-parameter model, https://x.com/AIatMeta/status/1932808881627148450
Figure’s robot works 60 minutes straight on warehouse tasks
Figure Robotics demonstrated their humanoid robot Helix performing logistics work continuously for an hour without human intervention. The robot now incorporates touch sensing and short-term memory, allowing it to handle various package types, scan barcodes with 95% accuracy, and adapt behaviors like patting down packages to ensure proper positioning. Performance improved 20% for package handling speed and 35% for barcode scanning compared to previous versions. The company attributes gains to scaling training data from 10 to 60 hours, which reduced processing time per package from 6.3 to 4.3 seconds. Traditional robotic programming approaches couldn’t achieve this level of adaptability across the variety of objects and subtle behaviors required for warehouse logistics.
Watch Helix’s neural network do 60 minutes of uninterrupted logistics work Helix now incorporates touch and short-term memory and it’s performance continuously improves over time https://x.com/Figure_robot/status/1931391490967928936
AI companies expand into military and defense contracts
OpenAI secured a $200 million Department of Defense contract to provide AI systems for national security and combat applications, while CEO Sam Altman indicated the company might help develop weapons systems. Anduril partnered with German defense company Rheinmetall to manufacture military drones for European markets. The US Army appointed executives from Palantir, Meta, and OpenAI as Lieutenant Colonels.
Anduril, Rheinmetall partner to build military drones for Europe | Reuters https://www.reuters.com/business/aerospace-defense/anduril-rheinmetall-partner-build-military-drones-europe-2025-06-18/
Never say never,”” OpenAI’s Sam Altman response on if the company would help the Pentagon develop a new weapons system. https://x.com/MsKristenTalman/status/1910483410545504474
OpenAI won a $200M U.S. DoD contract The one-year pilot will see OAI deliver frontier AI to address national security challenges in “”warfighting and enterprise domains”” This marks the lab’s foray into defense tech, while sticking to its usage policies https://x.com/rowancheung/status/1934881484722033140
US Army appoints Palantir, Meta, OpenAI execs as Lt. Colonels – The Grayzone https://thegrayzone.com/2025/06/18/palantir-execs-appointed-colonels/
Government agencies and national firms adopt AI for intelligence and administrative work
US officials report AI tools are accelerating analysis and decision-making across government operations. The UK government deployed a system called Extract that converts complex planning documents, including handwritten notes and blurry maps, into digital data within 40 seconds to help city planners make faster decisions. Meanwhile, Japan’s largest bank partnered with AI company Sakana to automate banking document creation, and Russia’s Sberbank announced plans for an AI system with advanced reasoning capabilities.
Tulsi Gabbard says AI is speeding up US intelligence work | AP News https://apnews.com/article/gabbard-trump-ai-amazon-intelligence-beca4c4e25581e52de5343244e995e78
Nvidia’s pitch for sovereign AI resonates with EU leaders | Reuters https://www.reuters.com/business/media-telecom/nvidias-pitch-sovereign-ai-resonates-with-eu-leaders-2025-06-16/
Sakana AI and MUFG sign agreement to automate creation of banking documents https://x.com/SakanaAILabs/status/1934264383510925732
Extract – a system built by the UK government, using our Gemini foundational model – will help council planners make faster decisions. 🚀 Using multimodal reasoning, it turns complex planning documents – even handwritten notes and blurry maps – into digital data in just 40s. https://x.com/GoogleDeepMind/status/1932032485254217799
Russia’s Sberbank plans to unveil LLM with reasoning capacity | Reuters https://www.reuters.com/business/finance/russias-sberbank-plans-unveil-llm-with-reasoning-capacity-2025-06-18/
Google demos AI generating user interfaces in real-time
Google demonstrated Gemini 2.5 Flash-Lite creating custom user interfaces instantly based on what appears on screen. The AI writes code for entirely new interfaces and their content in the time it takes to click a button, generating temporary interfaces tailored to specific tasks rather than using pre-built templates. The demonstration highlights how extremely fast AI models can enable new interaction patterns beyond traditional apps and websites. While the current version mimics conventional interface designs, the underlying concept suggests a shift toward AI creating personalized, task-specific interfaces on demand rather than users navigating through fixed menus and screens.
Cool demo of a GUI for LLMs! Obviously it has a bit silly feel of a “horseless carriage” in that it exactly replicates conventional UI in the new paradigm, but the high level idea is to generate a completely ephemeral UI on demand depending on the specific task at hand.”” / X https://x.com/karpathy/status/1935779463536755062
Hello Gemini 2.5 Flash-Lite! So fast, it codes *each screen* on the fly (Neural OS concept 👇). The frontier isn’t always about large models and beating benchmarks. In this case, a super fast & good model can unlock drastic use cases. Read more: https://x.com/OriolVinyalsML/status/1935005985070084197
Here’s how Gemini 2.5 Flash-Lite writes the code for a UI and its contents based solely on the context of what appears in the previous screen – all in the time it takes to click a button. 💻 ↓ https://x.com/GoogleDeepMind/status/1935719933075177764
Self-driving car industry accelerates toward commercial reality
The autonomous vehicle sector is hitting multiple milestones simultaneously across production, deployment, and research. Amazon’s Zoox opened its first major robotaxi manufacturing facility while Uber announced plans to launch self-driving taxis in London’s complex road network. Waymo published research proving that autonomous vehicle performance improves predictably when companies scale up data and computing power, following the same patterns seen in AI language models. Goldman Sachs analysts expect these developments to fundamentally reshape the auto insurance industry as liability shifts from human drivers to software systems, signaling that the technology is moving from experimental trials to mainstream commercial deployment.
Goldman Sees Autonomous Vehicles Transforming Insurance World – Bloomberg https://www.bloomberg.com/news/articles/2025-06-09/goldman-sees-autonomous-vehicles-transforming-insurance-world?srnd=undefined&embedded-checkout=true
Amazon’s Zoox opens its first major robotaxi production facility | TechCrunch https://techcrunch.com/2025/06/18/amazons-zoox-opens-its-first-major-robotaxi-production-facility/
New Insights for Scaling Laws in Autonomous Driving https://waymo.com/blog/2025/06/scaling-laws-in-autonomous-driving
Uber to launch self-driving taxis in London | Fortune Europe https://fortune.com/europe/2025/06/12/uber-launch-self-driving-taxis-london-complex-roads-andrew-macdonald/
Waymo shows that scaling data and compute lifts autonomous-vehicle forecasting and planning in a predictable way. • Similar to LLMs, motion forecasting quality also follows a power-law as a function of training compute. • Expanding data volume elevates model performance. • https://x.com/rohanpaul_ai/status/1934564747078508994
Salesforce launches AI agents to automate marketing campaigns
Salesforce released Marketing Cloud Next, a suite of AI tools that handles routine marketing tasks like building audience segments, writing email copy, and managing ad performance. The agents can work independently or assist marketers, taking directions like “build a campaign for high-value customers” and executing the entire process from strategy to deployment. The technology also transforms static marketing emails into interactive conversations where customers can ask questions and receive personalized offers in real time.
Apple refreshed its Apple Foundation Models (AFM) with new versions for on-device and server use, aiming to improve performance in tasks like image understanding and multilingual reasoning. The company also released a Foundation Models API for developers to integrate the https://x.com/DeepLearningAI/status/1936121879552537056
Microsoft launches Copilot Vision for Windows users nationwide
Microsoft made its AI assistant capable of viewing and analyzing what’s on your computer screen. Copilot Vision can now see your apps, photos, and webpages to provide real-time guidance, such as walking you through Adobe Photoshop features or answering questions about content you’re viewing. The feature works by letting users share specific browser windows or apps with the AI, similar to screen sharing in video calls. Unlike Microsoft’s controversial Recall feature that continuously captures screen activity, Copilot Vision only accesses apps when users specifically choose to share them through an opt-in process.
(BRAVEHEART) Microsoft’s new Copilot Vision can ‘see’ your apps on Windows | The Verge https://www.theverge.com/news/685963/microsoft-copilot-vision-windows-launch
OpenAI adds meeting and notes recording feature to ChatGPT – might be the end of a few startups – great for doctor visits
OpenAI rolled out a record mode for ChatGPT Team users on Mac computers that can capture meetings, brainstorming sessions, or voice notes. The feature automatically transcribes audio and extracts key points from conversations, then converts them into actionable follow-ups, plans, or code snippets. The tool addresses the common challenge of capturing and organizing ideas from verbal discussions, allowing teams to turn spoken conversations directly into structured outputs without manual note-taking.
Record mode is rolling out today in ChatGPT to Pro, Enterprise, and Edu users. Available on macOS desktop app.”” / X https://x.com/OpenAI/status/1935419375600926971
This is becoming a theme: if you are in a field where simple LLM wrappers have proliferated and are proven use cases with lots of adoptions (Copilot, Firefly, etc.) then the AI labs are likely to just toss that capability into their chatbot.”” / X https://x.com/emollick/status/1935409804681343057
People are using GPT and Claude to troll the Apple paper that said AI can’t reason
Researchers using AI assistance published a comprehensive analysis arguing that average humans often operate under an “illusion of thinking” where they believe their decisions stem from rational reasoning when they’re actually driven by psychological biases and social influence. The paper examines how fear responses, confirmation bias, and resistance to change constrain logical thought, while environmental factors like rigid education systems and digital overstimulation shape thinking patterns from childhood. The analysis draws parallels to recent AI research showing that reasoning models break down when problems become too complex, suggesting humans experience similar “accuracy collapse” when faced with novel challenges. Rather than engaging in deep analysis, people rely on mental shortcuts, learned scripts, and social mimicry that simulate genuine reasoning while maintaining the feeling of independent thought.
Claude Opus first authored it’s first paper, scientifically critiquing the claims of Apple researchers in “”Illusion of thinking”” paper 👀 https://x.com/reach_vb/status/1933635061561135312
Fun retort to Apple’s research that AI can’t think – ChatGPT – Shared Content https://chatgpt.com/s/dr_684a4e4aea4c819190bbe38f8dedc6a5
New medical AI model matches doctor performance on standard tests, runs on laptops
Intelligent Internet released II-Medical-8B-1706, an AI system designed to answer medical questions that performs as well as human physicians on healthcare benchmarks. The model outperformed Google’s larger MedGemma system while using 70% fewer parameters, making it efficient enough to run on computers with less than 8GB of memory. The company positions this as progress toward making medical knowledge more accessible, with the model achieving performance levels comparable to GPT-4 variants on medical evaluations. The smaller size allows the AI to run on standard consumer hardware rather than requiring expensive server infrastructure.
II-Medical-8B-1706 is our latest state of the art open medical model 💡 Outperforms the latest @Google MedGemma 27b model with 70% less parameters 🤏 Quantised GGUF weights, works on <8 Gb RAM 🚀 One more step to the universal health knowledge access that everyone deserves ⚕️ https://x.com/ii_posts/status/1934959488710094990
II-Medical-8B-1706, an 8B param health model reached GPT-4o/4.1/4.5 benchmarks, surpassing physicians. Intelligence is gradually ceasing to be human-exclusive. https://x.com/rohanpaul_ai/status/1935309527832056065
Intelligent Internet introduced II-Medical-8B-1706, an updated version of its open medical model Capable of running on <8GB RAM, the AI outperformed Google’s MedGemma 27B across benchmarks despite 70% fewer parameters https://x.com/rowancheung/status/1935247303524114645
OpenAI prepares safeguards for AI models with advanced biology capabilities
OpenAI is implementing multiple security measures as its AI models approach what the company calls “High” capability levels in biology. The models could accelerate drug discovery and vaccine development, but the same abilities that help scientists could potentially assist in creating biological weapons. The company has deployed always-on detection systems to block suspicious biology-related requests and uses expert “red teams” to test whether people can bypass safety measures. OpenAI plans to host a biodefense summit in July.
Preparing for future AI capabilities in biology | OpenAI https://openai.com/index/preparing-for-future-ai-capabilities-in-biology/
ByteDance releases Seedance 1.0 video generation model
ByteDance launched an AI system that creates high-quality videos from text prompts in under a minute. Seedance 1.0 generates 5-second videos at 1080p resolution in just 41.4 seconds, representing a 10-fold speed improvement over previous models through advanced optimization techniques. The system can handle consistent motion, complex instructions with multiple subjects, and coherent multi-shot sequences.
Seedance 1.0: Exploring the Boundaries of Video Generation Models – Publications – ByteDance Seed Team https://seed.bytedance.com/en/public_papers/seedance-1-0-exploring-the-boundaries-of-video-generation-models
So where did ByteDance Seed come from? The team was actually founded in 2023, but the “ByteDance Seed” branding didn’t become visible externally until around January 2025. Before that, their papers listed more generic ByteDance-affiliated names, without a consistent team”” / X https://x.com/arankomatsuzaki/status/1935603383014248764
Midjourney launches video generation from still images
Midjourney released its first video model that transforms static images into 5-second animated clips. Users can now click “Animate” on any Midjourney image to make it move, either automatically or by describing specific motions they want to see. The system offers high-motion settings for dynamic scenes with camera movement and low-motion for subtle ambient animations.The company positions this as a building block toward their ultimate goal of real-time interactive 3D worlds.
Introducing Our V1 Video Model https://www.midjourney.com/updates/introducing-our-v1-video-model
Introducing our V1 Video Model. It’s fun, easy, and beautiful. Available at 10$/month, it’s the first video model for *everyone* and it’s available now. https://x.com/midjourney/status/1935377193733079452
Latent space explorers — it is time. Midjourney video is live. https://x.com/bilawalsidhu/status/1935466320960618593
Midjourney launches its first video model, letting users turn images into short animated clips https://the-decoder.com/midjourney-launches-its-first-video-model-letting-users-turn-images-into-short-animated-clips/
Midjourney video is very fun just browsed through my gallery & hit “”animate”” on a few https://x.com/fabianstelzer/status/1935433791788478933
Betting platform airs $2000 AI-generated commercial during NBA Finals
Kalshi ran a surreal advertisement during Wednesday’s NBA Finals that featured AI-generated scenes of people swimming in pools of eggs, aliens drinking beer, and cowboys carrying chihuahuas. The 15-second spot promoting various betting markets was created entirely using Google’s Veo 3 text-to-video generator in just 2-3 days for $2,000 total production cost.The ad creator needed 300-400 AI generations to produce 15 usable clips, representing what he calls a 95% cost reduction compared to traditional advertising production. His process involved writing a script, using Gemini AI to generate shot prompts, then feeding those prompts into Veo 3 before editing the results in standard video software.
Here’s the $2,000 fully AI-generated ad that aired during the NBA Finals | The Verge https://www.theverge.com/news/686474/kalshi-ai-generated-ad-nba-finals-google-veo-3
Google partners with filmmakers to create AI-assisted short film
Google DeepMind collaborated with director Darren Aronofsky and filmmaker Eliza McNitt to produce “ANCESTRA,” a short film that combines live-action footage with scenes generated by Google’s Veo AI video model. The film, which premiered at the Tribeca Festival, required developing new AI capabilities, including motion matching to follow precise camera movements and the ability to seamlessly blend generated content with live-action footage. Google’s team of over 200 people worked with McNitt to create scenes that would have been difficult or expensive to produce using traditional visual effects, such as realistic newborn imagery and journeys through the human body.
Behind ANCESTRA: combining generative AI with live-action filmmaking https://blog.google/technology/google-deepmind/ancestra-behind-the-scenes/
More details emerge around Meta offerring $100 million signing bonuses to poach OpenAI talent
Last week, the buzz was that Meta is reportedly offering $100 million signing bonuses to recruit top employees from OpenAI. The bonuses are separate from annual compensation and represent an escalation in the AI talent war between major tech companies. Meta also attempted to acquire or hire Ilya Sutskever from his company Safe Superintelligence, though those efforts appear unsuccessful.
Founder, Alexandr Wang, Joins Meta to Work on AI Efforts | Scale https://scale.com/blog/scale-ai-announces-next-phase-of-company-evolution
META attempted to buy Ilya Sutskever’s Safe Superintelligence, and also attempted to hire him, according to reporting tonight by CNBC. https://x.com/AndrewCurran_/status/1935853120472612955
holy shit these guys are already centimillionaires so we’re not talking $100m signing bonus anymore first $1b signing bonus in history this is going to COST. Zuck clearly is in spend mode – if you think $100m bonuses are high – this is a guy who LOST $14b in 2022, $16b in https://x.com/swyx/status/1935468206019461470
Sam Altman says Meta is offering $100M signing bonuses to OpenAI staff. Not $100M annual compensation, just the signing bonus! He clowned Meta: “that’s not how you build a great culture.” Also said none of OpenAI’s best people are leaving. This AI talent war is crazy. https://x.com/Yuchenj_UW/status/1935116041866330378
Sam Altman says Meta offered $100 million bonuses to OpenAI employees | Reuters https://www.reuters.com/business/sam-altman-says-meta-offered-100-million-bonuses-openai-employees-2025-06-18/
OpenAI’s Codex merges 345,000 code changes in 35 days
OpenAI’s Codex system has automatically merged 345,000 pull requests on GitHub over the past month! The system appears to be handling substantial coding tasks autonomously, from bug fixes to feature implementations across thousands of software projects. OpenAI predicts fully autonomous coding agents will arrive within 18 months.
Earlier this year, I reached conviction that we’d see true agentic software engineers within 18mo, and I didn’t want to miss out on the fun. Excited to share that I’ve joined the @openai Codex team to work on this! We’re quickly growing the team in SF – hiring FS/product and”” / X https://x.com/majoshma/status/1934655710509641974
in the last 35 days, @OpenAI codex has merged 345,000 PRs on github. 345,000. AI is eating software engineering https://x.com/AnjneyMidha/status/1935865723328590229
Amazon AWS Developer creates AI security agent that learns and adapts
A cybersecurity researcher built Cyber-AutoAgent, an AI system that conducts security testing while learning from its attempts and creating custom tools on the fly. Unlike traditional AI that follows scripts, this agent combines memory storage, self-reflection, and dynamic tool creation to improve its performance during penetration testing. The system can remember past findings, analyze why certain attacks failed, and write new code tools when existing ones prove insufficient. The open-source project demonstrates how AI can move beyond simple command execution to exhibit expert-like behavior in specialized domains. During tests on intentionally vulnerable applications, the agent completed assessments in about 10 minutes by automatically discovering vulnerabilities, reflecting on failed attempts, and adapting its approach based on what it learned.
🚀 Autonomous Agents That Think, Remember, and Evolve An incredible project by Aaron Brown (from AWS team), showcasing how autonomous agents can move beyond simple tasks to reason, remember, and adapt – powered by Mem0, @awscloud Bedrock, and Strands Agents SDK 💡 From https://x.com/mem0ai/status/1932098203706613888
Sam Altman discusses robots timeline and Meta rivalry
OpenAI’s CEO appeared on his brother’s podcast to address competition and future plans. Altman claimed Meta is offering $100,000 bonuses to recruit OpenAI talent but none of the company’s top performers have accepted, suggesting OpenAI has a better path to artificial general intelligence and will ultimately be more valuable than Meta. Altman also outlined his vision for humanoid robots, predicting they’ll arrive within five to ten years once the mechanical engineering challenges are solved. He believes robots walking among people will feel more futuristic than current AI chatbots, calling humanoid robots “the dream” he wants to achieve.
Sam Altman cooked Meta on his brother’s podcast. Key takeaways: — Meta’s offering $100K bonus to poach OAI talent — None of OAI’s best have taken the offers — OAI has a better shot at AGI, will eventually be more valuable — Meta isn’t great at innovation https://x.com/rowancheung/status/1935247162150896061
Sam Altman | The Future of AI – YouTube https://www.youtube.com/watch?v=mZUG0pr5hBo
Sam Altman on AGI, GPT-5, and what’s next — the OpenAI Podcast Ep. 1 – YouTube https://www.youtube.com/watch?v=DB9mjd-65gw
Sam Altman on his brother’s podcast: ⦿ Humanoid robots are the dream. I really care about that. In five to ten years, we’ll have great humanoid robots. ⦿ It’s been a hard mechanical engineering challenge – even if we had a perfect brain right now, I don’t think we have the https://x.com/TheHumanoidHub/status/1935135554875834674
Tech giants accelerate AI chip development and infrastructure investments
AWS announced plans for an updated server CPU and a new AI training chip, while Apple revealed it’s exploring using AI to help design its own processors. SoftBank’s Masayoshi Son pitched a massive $1 trillion AI hub project in Arizona, and Nvidia partnered with Mistral AI to build AI cloud infrastructure using American-made chips. Google’s early investment in TPU chips in 2015 is now seen as particularly prescient, as it remains one of the few major tech companies not dependent on Nvidia for AI processing power.
AWS to introduce updated server CPU, new AI training chip – SiliconANGLE https://siliconangle.com/2025/06/18/aws-introduce-updated-server-cpu-new-ai-training-chip/
Apple eyes using AI to design its chips, technology executive says https://finance.yahoo.com/news/apple-eyes-using-ai-design-231623677.html
SoftBank’s Son pitches $1 trillion Arizona AI hub, Bloomberg News reports https://finance.yahoo.com/news/softbanks-son-pitches-1-trillion-053605238.html
google simply does not get enough credit for the TPU the amount of conviction & foresight it took to fund and build hardware specifically for for AI… in 2015… ten years later, literally everyone else is dependent on nvidia. except big G really a brilliant strategic move https://x.com/jxmnop/status/1934003515577303512
Today we’re announcing we’re going to build an AI cloud together with @MistralAI”” – Jensen @nvidia GTC today unveiling Mistral Compute This is a massive win for America and for open source Open models on US chips wil be the template for AI infrastructure buildouts globally https://x.com/AnjneyMidha/status/1932782589410226388
The reason why companies struggle to measure profit from generative AI despite widespread adoption
A new report reveals that nearly 80% of companies use generative AI, but the same percentage report no significant financial impact from their investments. The problem stems from an imbalance between widely deployed but low-impact tools like chatbots and employee assistants versus more transformative function-specific applications that remain stuck in pilot programs due to technical and organizational barriers. The report argues that AI agents, which can autonomously plan, remember, and integrate across business processes, offer a solution by transforming AI from a reactive tool into a proactive collaborator. However, companies need to redesign their workflows around agents rather than simply adding them to existing processes, requiring CEOs to shift from scattered experiments to strategic, cross-functional programs.
GenAI paradox: exploring AI use cases | McKinsey https://www.mckinsey.com/capabilities/quantumblack/our-insights/seizing-the-agentic-ai-advantage
McKinsey’s new report on AI agents shows the same mindset I see in many firms: a focus on making small, obsolete models do basic work (look at their suggested models!) rather than realizing that smarter models can do higher-end work (and those models are getting cheaper & better) https://x.com/emollick/status/1934669417494745559
Rewiring the way McKinsey works with Lilli, our generative AI platform | McKinsey Digital | McKinsey & Company https://www.mckinsey.com/capabilities/mckinsey-digital/how-we-help-clients/rewiring-the-way-mckinsey-works-with-lilli
Salesforce creates business benchmark showing AI agents struggle with real work
Salesforce developed CRMArena-Pro, a test that measures how well AI agents handle actual business tasks like sales, customer service, and pricing workflows. The benchmark uses realistic scenarios with multiple conversation turns and different customer personalities, plus tests whether agents can keep confidential information secure. Leading AI agents scored only 58% on single interactions and dropped to 35% when handling back-and-forth conversations. The results reveal a significant gap between AI capabilities and business needs, with agents showing almost no awareness of confidentiality requirements unless specifically prompted – which then hurt their task performance. While agents performed better on workflow execution (83% success), complex reasoning and maintaining security standards during multi-step conversations remain major challenges.
Interesting attempt by Salesforce to create a benchmark for realistic business tasks – we need more of these! Worth tracking over time (though I would love to see an contest, ARC-AGI style, to ask people to try to beat these benchmarks and see if they can with prompts & tools) https://x.com/emollick/status/1933175028033655241
Stanford study outlines worker preferences for AI automation
A Stanford study of 1,500 workers across 104 occupations found that 46% of workplace tasks receive positive ratings for AI automation, but workers often want different levels of AI involvement than what experts think is technically possible. Workers most want AI to handle repetitive, stressful tasks that free up time for higher-value work, while resisting automation in areas requiring human judgment or creative touch. The research reveals a mismatch between current AI investment and worker needs, with 41% of Y Combinator companies focusing on tasks workers either don’t want automated or consider low priority. Workers prefer equal partnership with AI systems rather than full automation, suggesting future workplace friction as AI capabilities advance beyond what employees find comfortable.
Future of Work with AI Agents https://futureofwork.saltlab.stanford.edu/
DOGE used flawed AI to identify VA contracts for cuts and pushed agencies to hire inexperienced contractor for sensitive data
The Department of Government Efficiency used artificial intelligence to review Veterans Affairs contracts, but the system contained significant technical problems that led to incorrect recommendations. The AI script, written by a DOGE employee, was programmed to flag contracts as cancelable if they weren’t “directly supporting patient care,” but it lacked the knowledge to make such determinations and mistakenly targeted essential services like internet connectivity. Independent experts who reviewed the code found the AI hallucinated contract values, incorrectly assigning $34 million price tags to agreements sometimes worth only thousands of dollars, and failed to analyze complete contract texts. The system used general-purpose AI models unsuitable for the specialized task of evaluating government contracts. DOGE leaders pressured government executives to hire a 21-year-old former Palantir intern and grant him access to personal data of every Social Security cardholder. Agency executives raised concerns that the individual lacked sufficient training to handle such sensitive information, but faced pressure to proceed with the arrangement despite their reservations about the security risks involved.
Inside the AI Tool Used by DOGE to Review Veterans Affairs Contracts — ProPublica https://www.propublica.org/article/inside-ai-tool-doge-veterans-affairs-contracts-sahil-lavingia
RT @MollyJongFast: “DOGE leaders pressured agency executives to hire a 21-year-old former intern at Palantir, a data analysis and technolog…”” / X https://x.com/zacharynado/status/1934766211738284114
Anthropic offers free courses on AI app development
Anthropic launched two free courses for developers and AI users. The first course teaches how to build AI applications using MCP (Model Context Protocol), showing developers how to connect AI agents to external data sources like GitHub, Google Docs, and local files. The second course covers AI fundamentals through their “AI Fluency: Framework & Foundations” program. The courses target different skill levels, with the MCP course focusing on technical implementation for developers while the AI Fluency course appears designed for broader understanding of AI concepts and applications.
AI Fluency: Frameworks and Foundations \ Anthropic https://www.anthropic.com/ai-fluency
Anthropic just dropped a free course on building AI Apps with MCP. Learn to connect AI Agents to external data sources like GitHub, Google Docs, local files using MCP. 100% free. https://x.com/Saboo_Shubham_/status/1929916710682783915
Investigation reveals OpenAI leadership and governance issues
A comprehensive report called “The OpenAI Files” documents alleged misconduct by CEO Sam Altman, including falsely listing himself as Y Combinator chairman in SEC filings for years and undisclosed financial conflicts of interest. The investigation claims Altman owns stakes in companies that later partnered with OpenAI, potentially netting him tens of millions in undisclosed gains, while telling Congress he holds no OpenAI equity. Multiple senior OpenAI executives, including CTO Mira Murati and co-founder Ilya Sutskever, reportedly told the board they were uncomfortable with Altman leading the company toward artificial general intelligence, citing what some called “gaslighting” and “psychological abuse.” The report also alleges OpenAI quietly modified its profit cap to allow unlimited growth, suffered an undisclosed security breach in 2023, and used restrictive agreements to prevent employees from criticizing the company even after leaving.
RT @robertwiblin: Huge repository of information about OpenAI and Altman just dropped — ‘The OpenAI Files’. There’s so much crazy shit in…”” / X https://x.com/NeelNanda5/status/1935642920662737290
The OpenAI Files https://www.openaifiles.org/
13 AI Visuals and Charts: Week Ending June 20, 2025
#CVPR2025 Picks #3 Alibaba just released VideoRefer-VideoLLaMA3 (2B & 7B video LLMs with A2.0 license!) These models can understand videos and segment objects, answer questions about them throughout the video at the same time 🤯 see it in action ⤵️ https://x.com/mervenoyann/status/1935739721772081336
Claude 4 Opus, o3-pro, o3 with image gen, and Gemini 2.5: make a midwit meme, make it awesome and insightful and funny”” Gemini struggles on the format, and only Claude gets close to actually pulling it off right, where the over-complication is in tension with the easy answer. https://x.com/emollick/status/1934479997294813682
AI stitching together multiple video feeds into one omniscient traffic god. This is what happens when cameras start talking to each other — mapping the trajectory of every vehicle and pedestrian seamlessly across cameras. Spatial intelligence is coming to a city near you. https://x.com/bilawalsidhu/status/1933346880336941297
How good is @runwayml Gen-4 References for visual effects? Here is a sample of what’s possible, all generated. https://x.com/c_valenzuelab/status/1934312626021949687
Woah. First person camera view being warped to a third person camera view using a unified flow and matching model. https://x.com/bilawalsidhu/status/1932975764992868427
I use FSD for over 2 hours of travel every day around the Philadelphia area and have never experienced an issue as bad as this one. Any ideas why it chose the completely wrong lane? Has anyone ever experienced anything like this? https://x.com/billykyle/status/1934300693235732801
Who did it best? Simple svg prompt, one-shot : r/singularity https://www.reddit.com/r/singularity/comments/1l486ji/who_did_it_best_simple_svg_prompt_oneshot/
No signs of an end to rapid gains in AI ability at ever-decreasing costs yet I did my best to update my chart to take into account the price drop in o3 & the new models released by Google. GPT-4 was 2.25 years ago,so its worth noting the trend when considering the future of AI. https://x.com/emollick/status/1935163667311480928
The progress of Gemini over the last year + https://x.com/OfficialLoganK/status/1935136191927501235
Veo 3 has digested the mother lode of ASMR content on YouTube making it an AI ASMR machine. This one got 3.1M likes and 12k comments in 3 days. Every popular YouTube format is about to get its impossible AI remix. https://x.com/bilawalsidhu/status/1934799544505483413
ANCESTRA by Eliza McNitt – YouTube https://www.youtube.com/watch?v=HEs9miwtwh4
The first film from our partnership with @primordialsoup_ – a storytelling venture founded by visionary director Darren Aronofsky – is debuting at @Tribeca. Directed by Eliza McNitt, ANCESTRA uses traditional filmmaking alongside Veo, our generative video model. Take a look ↓ https://x.com/GoogleDeepMind/status/1933549777192460771
Controllable and Expressive One-Shot Video Head Swapping https://humanaigc.github.io/SwapAnyHead/
Top 56 Links of The Week – Organized by Category
AGI
This recent report from the Joint California Policy Working Group on AI Frontier Models is an important step towards effective and balanced AI regulation in California. It builds meaningfully on the International AI Safety Report and offers a thoughtful framework for policymaking”” / X https://x.com/Yoshua_Bengio/status/1935479129899401243
GitHub is Leaking Trump’s Plans to ‘Accelerate’ AI Across Government https://www.404media.co/github-is-leaking-trumps-plans-to-accelerate-ai-across-government/
IBM aims to build the world’s first large-scale, error-corrected quantum computer by 2028 | MIT Technology Review https://www.technologyreview.com/2025/06/10/1118297/ibm-large-scale-error-corrected-quantum-computer-by-2028/
ARVR
Meta announces Oakley smart glasses that shoot 3K video | The Verge https://www.theverge.com/news/690133/meta-oakley-hstn-ai-glasses-price-date
🚀 Introducing Cosmos-Predict2! Our most powerful open video foundation model for Physical AI. Cosmos-Predict2 significantly improves upon Predict1 in visual quality, prompt alignment, and motion dynamics—outperforming popular open-source video foundation models. It’s openly https://x.com/qsh_zh/status/1933024567011995865
AgentsCopilots
Vibe coding is bad when you don’t know what you are doing (i.e., blind copy-pasting), or when you are a guru (AI is still too stupid to match your power). Somewhere in between, there is a sweet spot where vibe coding makes you a much happier coder.”” / X https://x.com/hyhieu226/status/1934113316965920950
RT @rohanpaul_ai: This is really BAD news of LLM’s coding skill. ☹️ The best Frontier LLM models achieve 0% on hard real-life Programming…”” / X https://x.com/sainingxie/status/1934994111536251361
Past the event horizon? OpenAI’s Sam Altman says so. New AI research backs him up https://tech.yahoo.com/ai/articles/past-event-horizon-openai-sam-163252151.html
This is a fantastic post from @AnthropicAI that should inform folks on what a production-grade multi-agent architecture looks like. There are three points worth calling out: 1️⃣ Not every use cases is suitable for multi-agents: “Further, some domains that require all agents to https://x.com/jerryjliu0/status/1934331886308110627
What if a livestream had two digital avatars—talking, reacting, and engaging in real time? Luo Yonghao, one of China’s top livestreamers, made his digital avatar debut on Baidu’s e-commerce platform. Powered by the ERNIE foundation model, the livestream was the first to feature https://x.com/Baidu_Inc/status/1934982099112751197
Starbucks’ new game plan to roll out AI chatbots at cafes could serve as a ‘litmus test’ for the industry, analyst says | Fortune https://fortune.com/2025/06/11/starbucks-ai-chatbot-green-dot-assist-turnaround-plan/
Just pulled an all nighter learning n8n from scratch… and i have to say this tool is insane. One app. So much power. I built a full automation with logic, AI, HTTP requests, Google Sheets, YouTube uploads… all visually. 🤯 Can’t believe I slept on this for so long. https://x.com/danielderedev/status/1923909340441850109
More Than 40% of Employees Are Using AI at Work, a New Poll Says – CNET https://www.cnet.com/tech/services-and-software/more-than-40-of-employees-are-using-ai-at-work-a-new-poll-says/
How Morgan Stanley Tackled One of Coding’s Toughest Problems https://www.msn.com/en-us/money/other/how-morgan-stanley-tackled-one-of-coding-s-toughest-problems/ar-AA1FZBGR
Google working on AI email tool that can ‘answer in your style’ | Artificial intelligence (AI) | The Guardian https://www.theguardian.com/technology/2025/jun/03/google-deepmind-ai-email-tool-answer-in-your-style
Andrej Karpathy: Software Is Changing (Again) Key learning points from this brilliant lecture from yesterday. 🚀 The Shifting Software Map For 70 years code flowed in one style, then neural networks arrived and rewrote large patches of logic. Karpathy divides eras into https://x.com/rohanpaul_ai/status/1935526340658094104
NickTikhonov/snap-ql: AI-powered Postgres Client https://github.com/NickTikhonov/snap-ql
more experiments on letting agents on @heyglif generate longer vids with Flux Ultra, Kling 2.1, MMAudio and automated stitching prompt was “”roman legionnaire travel log”” abundantly clear IMO that authorship will move from creating films to creating agents that create films https://x.com/fabianstelzer/status/1935038388782113197
Here’s how Gemini 2.5 Flash-Lite built a research prototype that can instantly transform large PDF files into interactive web apps – making it easier to summarize and understand dense information. 📄 Try it now in @Google AI Studio → https://x.com/GoogleDeepMind/status/1935005262097551811
I built an AI automation that turns product photos into engaging videos in seconds. Perfect for: -E-commerce stores -Social media content -Product demos -Marketing campaigns Want it? Just: Follow me RT + Comment “”n8n”” I’ll DM it to you https://x.com/samruddhi_mokal/status/1927968592416403946
AI spots heart disease warning signs in routine chest scans – Earth.com https://www.earth.com/news/ai-spots-heart-disease-warning-signs-in-routine-chest-scans/
how we automate our @beehiiv newsletter… 100% built in @gumloop_ai 🤖 step 1: pass in articles/YT videos/ websites step 2: get transcipts or scrape article step 3: use AI node to summarize in EXACT format step 4: COMBINE text with custom node step 5: output to sheet/RSS https://x.com/MakerThrive/status/1928481916719620174
The Browser Company launched Dia, an AI-first browser, in beta, with: —AI in the URL bar —Chatbot to analyze all tabs at once, draft emails, and answer with history context —Agentic skills for tasks like shopping & coding, with context from relevant tabs https://x.com/rowancheung/status/1933072129949347916
Anthropic
Anthropic CEO claims AI models hallucinate less than humans | TechCrunch https://techcrunch.com/2025/05/22/anthropic-ceo-claims-ai-models-hallucinate-less-than-humans/
Apple
We benchmarked Apple’s new On-Device model: trails most Gemma and Qwen on-device suitable models but still very useful GPQA Diamond performance trailed models that are suitable for on-device use such as the smaller Gemma models (3n E4B, 4B, 12B) and Qwen3 models (1.7B, 4B, 8B). https://x.com/ArtificialAnlys/status/1936141541023924503
Audio
Eleven v3 now supports Text to Speech in 41 new languages – bringing the total to over 70. This means you can now reach over 90% of the global population with ElevenLabs. https://x.com/elevenlabsio/status/1933557199294640171
Introducing ElevenLabs Conversational AI 2.0 – YouTube https://www.youtube.com/watch?v=TlclS4wLWgY
AutonomousVehicles
Waymo robotaxis are pushing into even more California cities | TechCrunch https://techcrunch.com/2025/06/17/waymo-robotaxis-are-pushing-into-even-more-california-cities/
Honda-backed Helm.ai unveils vision system for self-driving cars https://tech.yahoo.com/transportation/articles/honda-backed-helm-ai-unveils-100605030.html
BusinessAI
The Godfather of AI reveals which jobs are safest — and where ‘everybody’ will get replaced https://www.yahoo.com/news/godfather-ai-reveals-jobs-safest-163407655.html
Getty CEO: Stability AI lawsuit doesn’t cover industry mass theft https://www.cnbc.com/2025/05/28/getty-ceo-stability-ai-lawsuit-doesnt-cover-industry-mass-theft.html
“I find the story of AI and radiology fascinating. Of course, Hinton’s prediction was wrong* and tech advances don’t automatically and straightforwardly cause job replacement — that’s not the interesting part. Radiology has embraced AI enthusiastically, and the labor force is https://x.com/random_walker/status/1935679764192256328
Meta in Talks to Hire AI Investors Friedman and Gross, Partially Buy Out Their Venture Fund — The Information https://www.theinformation.com/articles/meta-talks-hire-former-github-ceo-nat-friedman-daniel-gross-join-ai-efforts
Elon Musk’s xAI in Talks to Raise $4.3 Billion in Equity Funding – Bloomberg https://www.bloomberg.com/news/articles/2025-06-17/musk-s-xai-in-talks-to-raise-4-3-billion-in-equity-funding?embedded-checkout=true
EducationAI
Ohio State announces every student will use AI in class | NBC4 WCMH-TV https://www.nbc4i.com/news/local-news/ohio-state-university/ohio-state-announces-every-student-will-use-ai-in-class/
A useful piece on criticizing AI: “its all PR” and “they are just parrots” are increasingly dead ends in a world where AI clearly can do effectively novel & important tasks. AI calls for robust criticism, but that criticism needs to be more grounded in the current state of LLMs.”” / X https://x.com/emollick/status/1935020901172387979
Introducing “”Building with Llama 4.”” This short course is created with @Meta @AIatMeta, and taught by @asangani7, Director of Partner Engineering for Meta’s AI team. Meta’s new Llama 4 has added three new models and introduced the Mixture-of-Experts (MoE) architecture to its https://x.com/AndrewYNg/status/1935350552692658202
Your Brain on ChatGPT: Accumulation of Cognitive Debt when Using an AI Assistant for Essay Writing Task https://arxiv.org/pdf/2506.08872v1
EthicsLegalSecurity
One of the most effective things the U.S. or any other nation can do to ensure its competitiveness in AI is to welcome high-skilled immigration and international students who have the potential to become high-skilled. For centuries, the U.S. has welcomed immigrants, and this”” / X https://x.com/AndrewYNg/status/1935741989204770837
understanding_the_impacts_of_generative_ai_use_on_children_-_wp1_report.pdf https://www.turing.ac.uk/sites/default/files/2025-06/understanding_the_impacts_of_generative_ai_use_on_children_-_wp1_report.pdf
Getty Images CEO warns it can’t afford to fight every AI copyright case | TechSpot https://www.techspot.com/news/108138-getty-images-ceo-warns-cant-afford-fight-every.html
A big AI question is why, as LLMs get bigger, their values seem to increasingly converge on the same preferences, and this holds for Musk’s Grok and China’s DeepSeek, too. “These findings suggest that value systems emerge in LLMs in a meaningful sense, with broad implications” https://x.com/emollick/status/1934278025685901374
We just shipped video FPS support in the Gemini API, so you can dynamically customize how many frames per second you want the model to see, unlocking lots of interesting new video use cases! 📹 https://x.com/OfficialLoganK/status/1935444350374125983
OpenAI
Whenever we release something, you all ask about custom GPTs 🥲 Good news! You can now set a recommended model when you create a custom GPT, and paid users can use our full range of models* when using a custom GPT. (* = GPTs with custom actions are limited to 4o and 4.1 for now)”” / X https://x.com/kevinweil/status/1935722240009437635
When people are delighted… it when humanity loses to AI…..There’s a handful of personal use cases that make me really bullish on hyper personalized utility. One of them is using ChatGPT as a personal running coach. Fed it all my run stats going back years, and said hey I have a race on this date, I wanna run x pace and keep my HR”” / X https://x.com/raizamrtn/status/1935781113513091107
Toward understanding and preventing misalignment generalization | OpenAI https://openai.com/index/emergent-misalignment/
Barbie is getting an AI upgrade as Mattel partners with OpenAI to develop smart toys using their technology, marking a new chapter for the iconic brand. https://x.com/Adweek/status/1933207258365579320
My beef with this is they don’t report the model number. LOL. They Asked ChatGPT Questions. The Answers Sent Them Spiraling. – The New York Times https://www.nytimes.com/2025/06/13/technology/chatgpt-ai-chatbots-conspiracies.html
A Cheeky Pint with OpenAI cofounder Greg Brockman – YouTube https://www.youtube.com/watch?v=E6hCFDfkijU
Perplexity
Howard Marks now writes his memo with Perplexity. “I have simplified the format and added emphasis, but haven’t changed a word. What follows below is pretty close to what I would have produced in an hour or two”. Warren Buffett famously said, “When I see memos from Howard Marks https://x.com/AravSrinivas/status/1935913410119844130
Make the Internet delightful again @PerplexityComet https://x.com/AravSrinivas/status/1936137070134853875
Publishing
Meet Dia. Now available for Arc members. https://x.com/diabrowser/status/1932800009990517190
Browsers are the perfect place for hybrid AI inference, combining the power of cloud AI with the versatility and privacy of local models. ⚡️ Dia already has great WebGPU support, so I’m looking forward to seeing more on-device models being integrated directly into the browser https://x.com/xenovacom/status/1935092938922758481
Robotics
BREAKING at #BAAI Conference 2025: BAAI unveils RoboOS 2.0 (cross-embodied brain and cerebrum collaboration framework) & RoboBrain 2.0! Outperforms mainstream embodied AI models -World’s strongest open-source embodied brain model!#Robotics #EmbodiedAI #OpenSource https://x.com/BAAIBeijing/status/1931916124473499676
1X World Model Scaling Evaluation for Robots https://x.com/1x_tech/status/1934634700758520053
TechPapers
RT @NielsRogge: “”Hugging Face is basically the equivalent of Github in the era of software 2.0″” – Karpathy, 2025, colorized https://x.com/reach_vb/status/1935970251004313788





Leave a Reply