Bringing Agentic Workflows into Inflection for Enterprise

 LlamaIndex and Box are two great tools that work great together!

🔍 Use Box AI prompt tool to query documents without downloading

📊 Extract structured data from unstructured content with Box AI extract

🔗 Seamlessly integrate Box AI with LlamaIndex agents

Use Prolog to improve LLM’s reasoning – Shchegrikovich LLM

“Our customers have been getting a ton of great results using our email writer, so we decided to double down on this feature and build the best AI email writer on the internet. It’s the AI email writer for people who want to be more efficient but hates AI email slop. Easily 

“this is going to sound unhinged but the hype around “computer use” is indicative of profound misunderstanding of the strengths and weaknesses of language models” / X

“Can we teach small, local LMs to *use* large, remote LMs as *tools*? Papillon with @Sylvia_Sparkle & team shows that this can be very consequential for privacy. Using DSPy optimizers, we can teach Llama3-8B to reach 86% of frontier LLMs’ quality while hiding your private data!” / X

“Excited to share that @crewAIInc raised $18 million in funding, with our series A led by @insightpartners, with @Boldstartvc leading our inception round. We’re also thrilled to welcome @BlitzVentures, Earl Grey Capital(@amitvasudev_), and top AI leaders like @AndrewYNg and” / X

“I 💙 CrewAI and I am thrilled to announce a new partnership and integration in watsonx with CrewAI! In September, after exchanging a few DMs on LinkedIn, I had the pleasure of meeting @joaomdmoura , the founder and creator of CrewAI, in San Francisco. Over coffee, we quickly 

“YC S23’s @AutotabAI is an AI worker that uses a computer like you—so anything you can do, it can do. Engineered for hallucination-proof reliability, it executes repetitive tasks fully autonomously. 

“Scripted dialogue also known as “barks” in game dev has been the way NPCs utter dialogues in virtual worlds. Now with AI-powered characters, these characters can have a generated conversation based on environment cues, character backgrounds, or even prompted topics. 

Polish radio station replaces journalists with AI ‘presenters’ | CNN Business

Microsoft introduces ‘AI employees’ that can handle client queries | Microsoft | The Guardian

“Here’s a breakdown of most of the top model scores on aider’s code editing benchmark: 84% Claude 3.5 Sonnet 10/22 80% o1-preview 77% Claude 3.5 Sonnet 06/20 72% DeepSeek V2.5 72% GPT-4o 08/06 71% o1-mini 68% Claude 3 Opus” / X

“We are heartbroken by the tragic loss of one of our users and want to express our deepest condolences to the family. As a company, we take the safety of our users very seriously and we are continuing to add new safety features that you can read about here:” / X

Can a Chatbot Named Daenerys Targaryen Be Blamed for a Teen’s Suicide? – The New York Times

Gartner: 2025 will see the rise of AI agents (and other top trends) | VentureBeat

“The most advanced Muti Agent System for GTM is here. Consolidate all your GTM tools into a single, multi-agent GTM stack! With AI SDR, AI-powered Calling Agents, Prospecting Agents, Market Research Agents, an agent-verified lead database, and a modern AI-native CRM all in one 

“📹Building Useful Apps With Small, Local LLMs Our very own @Hacubu gave a talk at CascadiaJS on using LangGraph.js to build with smaller LLMs, and some of the benefits to designing apps to work with OSS models 

“GenAI Agents: Comprehensive Repository for Development and Implementation 🚀 This repository by Nir Diamant is a fantastic resource for getting started with agents Contains implementations of the most common agent architectures in LangGraph 

“🍀Build a JavaScript AI Agent With LangGraph.js and MongoDB What You’ll Learn: ⭐ How to use LangGraph.js ⭐ How to integrate MongoDB Atlas ⭐ How to build an AI agent that can manage conversations, look up data, and persist state across sessions 

“🎉🎂 LangChain turns 2! Two years ago, we launched a Python package called `langchain` to help developers build LLM applications that could reason. Today, we’re still chasing that same mission — and as the generative AI ecosystem has evolved, we’ve evolved too. LangChain has 

OSWorld: Benchmarking Multimodal Agents for Open-Ended Tasks in Real Computer Environments

“Perplexity Pro is transitioning to a reasoning powered search agent for harder queries that involve several minutes of browsing and workflows. Check it out! It’s still in beta and has some issues but we’re going to keep making it better! 

“Lots of people and companies are sleeping on the power of synthetic data. LLMs are ridiculously good at generating synthetic data but it’s not straightforward plus we need more novel and complex data for not only improving LLMs but also systems built on LLMs (agents, RAG, etc.) 

“Dropping the first few videos on my knowledge assistant video series 👇 Step 1: Figure out how to define an agentic workflow on top of your standard RAG endpoints that can use LLMs to reason before your retrieval layer and afterwards. This lets you build more sophisticated 

“🚀 Say hello to the xAI Agent 🤖🧠 Built a finance agent, data analyst and web search agent using the grok-beta model. Waiting for native twitter search + fun mode and this will be 🔥 Try it yourself: 

“We’re working on advanced autonomous agents! Every human will have an AI agent capable of performing complex tasks on their behalf. Join the Starfleet team @xai to help us build the future: 

“🚀 Say hello to the xAI Agent 🤖🧠 Built a finance agent, data analyst and web search agent using the grok-beta model. Waiting for native twitter search + fun mode and this will be 🔥 Try it yourself: 

Asana Announces AI Studio: No-Code Builder for Designing and Deploying AI Agents in Critical Workflows • Asana, Inc.

“New short course: Practical Multi AI Agents and Advanced Use Cases with crewAI. Learn to build and deploy advanced agent-based systems in real applications in this course, created with @crewAIInc and taught by its founder, @joaomdmoura! (Disclosure: I’ve made a small seed 

“Agentic Information Retrieval This paper provides a good introduction to agentic information retrieval, which is shaped by the capabilities of LLM agents. I’ve been developing with this paradigm recently and it does offer lots of interesting ways to optimize retrieval systems. 

“This is my favorite search feature. Eliminates AI agent hallucinations. How? • lets agent request specific data points • api returns only those data points Result: your agent doesn’t have to parse large JSON, eliminating hallucinations. 

“Simple text-to-SQL is trivial, actual enterprise-grade text-to-SQL is hard. There should be a lot more tutorials like the one from @kiennt_ on creating a proper SQL agent. You need to first map out and index your entire data catalog, and figure out how to retrieve from it 

CrewAI now lets you build fleets of enterprise AI agents | VentureBeat

Coinbase introduces ‘Based Agent’ for creating AI agents in 3 minutes

“Fantastic Research. DIAMOND (DIffusion As a Model Of eNvironment Dreams), a reinforcement learning agent trained in a diffusion world model. The model is generating image frames, rather than obscure “render”. During training model is fed footage of someone playing the game 

“AI agents will soon automate most repetitive tasks, freeing us for more impactful, creative work. Jobs requiring intelligence but following repetitive patterns will change. Companies and individuals embracing this shift will become more efficient and excel. Focus on results, 

Anthropic

“I got to play with the new Claude model that controls a mouse and keyboard last week. Full post shortly, but I had it play Paperclip Clicker (of course) and it did well over a hundred moves executing a coherent strategy without any intervention. Agents start to come into view. 

“Claude 3.5 Sonnet’s current ability to use computers is imperfect. Some actions that people perform effortlessly—scrolling, dragging, zooming—currently present challenges. So we encourage exploration with low-risk tasks. We expect this to rapidly improve in the coming months.” / X

“Introducing an upgraded Claude 3.5 Sonnet, and a new model, Claude 3.5 Haiku. We’re also introducing a new capability in beta: computer use. Developers can now direct Claude to use computers the way people do—by looking at a screen, moving a cursor, clicking, and typing text. 

“Playing with Claude Computer Use is very worthwhile. It’s obvious that its something that’ll be used in the future, much like when you first try ChatGPT or amazing tech like AirPods. BUT, it’s clear its integration will take some serious time. Here’s an example web task, 

Introducing computer use, a new Claude 3.5 Sonnet, and Claude 3.5 Haiku \ Anthropic

Claude | Computer use for coding – YouTube

Claude | Computer use for automating operations – YouTube

“Claude’s “computer use” beta is wild because you don’t need to make custom tools for LLMs to use — automation is about to look a lot more like screen recording a task/workflow involving any desktop apps, and asking Claude to take control and do it for you. 

“Anthropic’s computer use can operate mobile devices including iOS, Android, and mobile browsers 📱 Here it is ordering me an Uber and posting for me on X. 

“Just got our first AI-ordered pizza with Lindy + Claude computer use 🙂 

Computer use (beta) – Anthropic

“We’re trying something fundamentally new. Instead of making specific tools to help Claude complete individual tasks, we’re teaching it general computer skills—allowing it to use a wide range of standard tools and software programs designed for people. 

Introducing the analysis tool in Claude.ai \ Anthropic

anthropic-quickstarts/computer-use-demo at main · anthropics/anthropic-quickstarts · GitHub

“One of those “AI feels like a superpower” moments. I went to an old tweet about a diagram of city streets by entropy, and pasted the scientific paper and image into Claude and asked it to create code to replicate it. It built the code in one shot, even replicated color scheme 

Claude | Computer use for orchestrating tasks – YouTube

“The ability of multimodal AI to “understand” images is underrated. I just took these. Given the first photo Claude guesses where I am. Given the second it identifies the type of plane. These aren’t obvious. 

“Anthropic computer use API + iPhone mirroring to a Mac = AI controlled phone. Watch Claude control my phone and successfully look up stats in my Sports app. I even got it to play a game in the Chess app against another AI – pretty crazy. And this is the worst it’ll ever be. 

“The new Claude 3.5 Sonnet is the first frontier AI model to offer computer use in public beta. While groundbreaking, computer use is still experimental—at times error-prone. We’re releasing it early for feedback from developers. 

“We’ve built an API that allows Claude to perceive and interact with computer interfaces. This API enables Claude to translate prompts into computer commands. Developers can use it to automate repetitive tasks, conduct testing and QA, and perform open-ended research. 

Initial explorations of Anthropic’s new Computer Use capability

Show HN: Agent.exe, a cross-platform app to let 3.5 Sonnet control your machine | Hacker News

“The most significant gains are in coding. The new 3.5 Sonnet sets a new state-of-the-art on SWE-bench Verified with a score of 49% (using no complex scaffolding)—besting all models including reasoning models like OpenAI o1-preview and specialized models for agentic coding. 

“Computer use API We’ve built an API that allows Claude to perceive and interact with computer interfaces. You feed in a screenshot to Claude, and Claude returns the next action to take on the computer (e.g. move mouse, click, type text, etc). 

“I can’t tell you the last time I was so excited to see a new AI capability in action. We plugged in Claude computer use in @Replit Agent as a human feedback replacement. And… it just works! I feel it won’t take long until our agent will become fully autonomous. 

“Claude 3.5 Haiku 3.5 Haiku replaces 3.0 Haiku as our fastest and least expensive model. It outperforms many state-of-the-art models on coding tasks—including the original Claude 3.5 Sonnet and GPT-4o. 3.5 Haiku will be made available in the coming weeks. 

“The new Claude 3.5 Sonnet has *insane* capabilities when used as a Minecraft agent. It’s powered by a project called Mindcraft. Running this code allows you to spawn AI bots that will follow your instructions, build, and play the game. Here’s how to set it up in <15min. 

“🚨Anthropic just released the most amazing AI technology I’ve ever used I’m not kidding AI agents are here and you can now build your own personal army of AI’s that will do work for you Here is your demo and complete beginner’s guide: (trust me, you want to bookmark this) 

“New @AnthropicAI Computer Use feels surreal. But don’t take my word for it. We made a template on Replit for you to try. Watch me fork the template, ask the agent to go to YouTube, find a video, and even skip the ads — all in a few minutes. 

Anthropic announces AI agents for complex tasks, racing OpenAI

LangChain

“🚀Launching CoAgents Public Beta 🪁: Everything you need to build Agent-Native applications, powered by LangGraph & CopilotKit. CoAgents enables in-app agents with: – Agenetic generative UI ✨ – Shared state (between agent <-> application) – Streaming intermediate agent state – 

Microsoft Copilot News

“Copilot is the UI for AI, and with Copilot Studio, customers can easily create, manage, and connect agents to Copilot. Today we announced new autonomous agent capabilities across Copilot Studio and Dynamics 365 to help scale the impact of every individual, team, and business” / X

“Microsoft rebranding Copilot as ‘agents’? That’s panic mode. Let’s be real—Copilot’s a flop because Microsoft lacks the data, metadata, and enterprise security models to create real corporate intelligence. That is why Copilot is inaccurate, spills corporate data, and forces 

Microsoft to allow autonomous AI agent development next month

OpenAI 

“OAI seems to be leaning hard into multi agents for its next act just heard Noam Brown at TED AI talk about how his career has basically been building agents for games, lost faith for a bit to build o1, and now he is starting up the multiagent team at @openai. el presidente https://twitter.com/swyx/status/1849239462406148514

Leave a Reply

Trending

Discover more from Ethan B. Holland

Subscribe now to keep reading and get access to the full archive.

Continue reading