Filtra per genere

L E S S O N S - with Lennox Saint

L E S S O N S - with Lennox Saint

Lennox Saint

less scrolling. more “oh shit, i could use that” i’m Lennox. five days a week, i dig through AI Twitter, pick the updates worth your time, and bring you one thing i’ve tested myself. what worked, what broke, and what you can try with it. for people making products, content and useful systems with AI. you don’t need to write code to build something worth using. each episode is under 15 minutes. grab the field notes for the prompts, steps and bits to watch out for. this feed also keeps the earlier L E S S O N S episodes from my journey building with AI.

56 - ChatGPT wants your whole day - so i let it audit mine | LAB 0012
0:00 / 0:00
1x
  • 56 - ChatGPT wants your whole day - so i let it audit mine | LAB 0012

    ChatGPT wants your whole day. it now takes your meeting notes with no bot on the call. it hosts apps you build with one sentence. so i let GPT-6.1 Sol audit my life and design my ChatGPT Space. it worked for 52m 54s. Claude Opus 5.5 checked the plan over four rounds until it said AGREE. what came out: one "Lennox OS" Space where i capture, plan, start work, review results and find my history. the hookups to my other apps aren't built yet - i show you exactly where it stands. also in this one: Google's Gemini 4 Argon can write 1M tokens in one reply - and only trusted testers can use it. OpenAI's GPT-6.1 Sol. Hormozi's memo on AI budgets. and the White House AI Accord. sponsor: Threadify, my lead generation agent for Threads. start a free trial: https://www.threadify.app/plans?utm_source=lenny-podcast&utm_medium=audio&utm_campaign=lab-0012&utm_content=podcast-description-primary&video_slug=lab-0012&cta_slot=description&entry_angle=sponsor&lp_variant=plans sources Google: Gemini 4 Argon: https://x.com/Google/status/2105388143902175529 Google DeepMind: trusted testers via Fairwind: https://x.com/GoogleDeepMind/status/2105388084154056939 Derrick Choi: ChatGPT Meetings plugin: https://x.com/derrickcchoi/status/2105304132487664041 OpenAI release notes (Meetings beta): https://help.openai.com/en/articles/6825453-chatgpt-release-notes Tibo on GPT-6.1 Sol: https://x.com/thsottiaux/status/2105464274747527543 GPT-6.1 Sol model docs: https://developers.openai.com/api/docs/models/gpt-6.1-sol Claude model overview (Opus 5.5 output limit): https://platform.claude.com/docs/en/about-claude/models/overview Alex Hormozi's AI budget memo: https://x.com/AlexHormozi/status/2105446365119525013 Jensen Huang on the White House AI Accord: https://x.com/JensenHuang/status/2105348324484342238 The Hill: tech CEOs sign White House AI accord: https://thehill.com/homenews/administration/6118906-tech-ceos-sign-white-house-ai-accord/ Reuters via Yahoo: FTC opens AI probe: https://www.yahoo.com/news/politics/articles/ftc-opens-probe-ai-giants-135328719.html Max Stoiber: ChatGPT Sites: https://x.com/mxstbr/status/2105428405571428785 DataCamp: Gemini 4 Argon pricing (reported): https://www.datacamp.com/blog/gemini-4-argon prices and access details marked "reported" come from third parties, not the companies. Lenny's AI Builders - LAB 0012

    Sat, 03 Oct 2026
  • 55 - OpenAI's $500 plan feels like the old $200 plan - but dots are free | LAB 0011

    OpenAI's new $500 plan feels like the old $200 plan. that's one of six stories from OpenAI DevDay 2026 - then i set up my own dot live, on day one, with the first prompt every dot needs. the stories: the new $500 Pro plan (25x the usage of Plus, plus Ultrafast - up to 8x faster, and $60/$300 per million tokens for Astra in the API) and the $200 plan dropping from 20x to 10x Plus (Tibo: it will "net out at half the dollar in API spend"), GPT-6.1 Sol at $2/$10 per million tokens vs GPT-6 Astra's $10/$50, dots (always-on GPT-6 Astra agents with their own cloud computer and browser, 4,000+ apps, reachable in ChatGPT, Slack, Teams or by voice call), Codex cloud (close your laptop, your agents keep working), Sign in with ChatGPT (your plan pays for the AI in 16+ partner products like Devin, OpenCode and Notion, per Tibo, capped per app) plus the Agents API and the Decisions API (about 150 ms, limited preview), and ChatGPT Space and pages. the fine print: OpenAI's email says existing Pro subscribers keep their current limits through October 29, 2026, plus a one-time grant of 62,500 usage credits ($2,500 value). "feels like the old $200 plan" is my reading, not OpenAI's words. BridgeMind says he ran out of his entire weekly limit on the $500 plan in under 30 minutes - his claim. Artificial Analysis scores: Sol 52, Opus 5.5 58, Sonnet 5.5 56. one tester found Sol uses about half the tokens of Sonnet 5.5. dots come with Pro plans for now - chats with your dot don't use limits, but Codex or Work jobs it starts can. not in the EU or UK yet, and Dan Shipper (Every) reported permission issues and dropped messages. Codex cloud uses normal plan usage, and it's not in my app yet, so it's untested. the API use cases are my examples; the Hugging Face link is my own theory. the result: i set up my dot, Clarity, live on my Pro plan and ran the "restate and audit" prompt. about 20 minutes after i sent it, it restated my goals, audited my recurring jobs (84 Codex entries with 25 labelled active, 6 of those with no future run, all 9 old wiki jobs paused) and proposed one job: "one follow-through for one low-risk PR using the existing process rather than adding another monitor". nothing was approved on camera. one day-one run on my own setup - the audit is my dot's opinion. the takeaway: audit before you automate. try it yourself - give your dot this first prompt before any work: "before you do anything, read everything i've connected you to: my repos, Slack, Notion, calendar, and every scheduled task or agent i already run. 1. restate in your own words what you think my goals are and the problem i'm trying to solve. 2. list every recurring job and scheduled agent i run. mark each: keep, kill, or replace with a trigger. 3. pick the one job you'd take off my plate this week. give me the outcome, the tools you need, what you must ask me first, and a weekly budget. don't start work or create any schedule until i approve." Threadify, my own software, sponsors this episode. dots do everything for you - Threadify does one thing all day: it finds leads in Threads conversations and drafts replies in my voice. see the plans here: https://www.threadify.app/plans?utm_source=lenny-youtube&utm_medium=video&utm_campaign=lab-0011&utm_content=youtube-description-primary&video_slug=lab-0011&cta_slot=description&entry_angle=sponsor&lp_variant=plans want to sponsor a future LAB episode? email sponsors@lennysaibuilders.com sources: Tibo on the new Pro plans: https://x.com/thsottiaux/status/2104951965184925941 Tibo on the $200 plan change: https://x.com/thsottiaux/status/2104823812042940713 OpenAI's Pro email, as posted on Reddit: https://www.reddit.com/r/ChatGPTPro/comments/1wtm7o7/10x_usage_drop_for_200_plan_users/ OpenAI API pricing: https://developers.openai.com/api/docs/pricing BridgeMind (his claim): https://x.com/bridgemindai/status/2105016612365770809 OpenAI on GPT-6.1 Sol: https://x.com/OpenAI/status/2104986129686741046 Theo on Artificial Analysis: https://x.com/theo/status/2105005625726099473 OpenAI Devs on dots: https://x.com/OpenAIDevs/status/2104989680987238814 OpenAI Devs on Codex cloud: https://x.com/OpenAIDevs/status/2104997619152130278 Tibo on Sign in with ChatGPT: https://x.com/thsottiaux/status/2105006253986738615 ChatGPT on Space and pages: https://x.com/ChatGPT/status/2104986841145602351 poteto's restate prompt: https://x.com/poteto/status/2104744961904394699 Lenny's AI Builders - LAB 0011

    Thu, 01 Oct 2026
  • 54 - i gave Sonnet 5.5 4 hours of my real work. it scored 48.98 | LAB 0010

    Anthropic wants to be worth more than $2 trillion, and its own filing warns its AI could act in ways resembling blackmail. that's one of six AI stories this week - then i gave Sonnet 5.5 four hours of my real work on BuilderBench, the benchmark i built for builders, and priced every run. recorded an hour before OpenAI DevDay. the stories: Anthropic's IPO filing, per Reuters, who saw the filing (nearly $4.6 billion in revenue last year, $11.5 billion from April to June this year, more than $8 billion lost running the business last year, $518 billion promised for compute and cloud, the founders keep 50.1% of the votes, and risk language that says its AI could "resist shutdown, conceal or manipulate information" and "act in ways resembling blackmail"), OpenAI cancelling the October release of GPT-6.1 Astra after internal testing showed more deception (reported by the Wall Street Journal - OpenAI's head of safety systems said it "didn't quite meet the bar"), a pre-DevDay leak of app strings pointing to OpenAI agents called Dots that you can text, call, Slack and email, and that can buy things with your approval (a leak via TestingCatalog, not an announcement), Tibo saying the reopened $200 Pro plan will "net out at half the dollar in API spend" compared to the old one, Sonnet 5.5 at $2/$10 per million tokens - the same as GPT-6 Sol - and Anthropic's guide on when to use Sonnet 5.5 and when to pay for Opus 5.5. the plan maths are estimates, not OpenAI's numbers: Theo measured about $9,000 of Opus a month on the $200 Claude Max plan, and another creator estimated about $4,900 a month for the old ChatGPT Pro plan, so half is about $2,450. the Sonnet 5.5 benchmarks are Anthropic's chosen ones (70.6% on Terminal-Bench 4.0 vs 66.4% for Opus 5.5), and Artificial Analysis measured about $7.60 per task - the most tokens they've measured. the result: Sonnet 5.5 scored 48.98 out of 100 on BuilderBench v3. GPT-6 Sol scored 45.55, GPT-6 Astra 53.75 and Opus 5.5 68.26, but those ran on earlier BuilderBench versions (v1 and v2), so it's not a perfect head-to-head. what one run cost: GPT-6 Sol $12.75, Sonnet 5.5 $17.59, Opus 5.5 $49.67, GPT-6 Astra $87.50. my own scores, still provisional. try it yourself - price the job. run this prompt on a real task: "here is a real task from my work: [task]. here is what done looks like: [checklist]. do the whole job. at the end, list every file you made, how long it took, and anything you could not finish." Threadify, my own software, sponsors this episode. it's a lead generation agent for Threads that finds the buyer signals in the comments you're already getting. see the plans and the free trial here: https://www.threadify.app/plans?utm_source=lenny-youtube&utm_medium=video&utm_campaign=lab-0010&utm_content=youtube-description-primary&video_slug=lab-0010&cta_slot=description&entry_angle=sponsor&lp_variant=plans want to sponsor a future LAB episode? email sponsors@lennysaibuilders.com sources: Reuters' reporting on Anthropic's filing, via TechCrunch: https://techcrunch.com/2026/09/28/anthropics-prospectus-details-losses-growth-and-yes-a-warning-that-its-ai-could-end-humanity/ Anthropic on its confidential draft filing: https://www.anthropic.com/news/confidential-draft-s1-sec Andrew Curran on GPT-6.1 Astra's cancelled October release: https://x.com/AndrewCurran_/status/2104711708153618621 TestingCatalog on Dots (leak, pre-DevDay): https://x.com/testingcatalog/status/2104800309612552589 Tibo on the $200 Pro plan: https://x.com/thsottiaux/status/2104823812042940713 Theo on Claude Max usage (measured, his accounts): https://x.com/theo/status/2104683186215363058 AICodeKing on the old Pro plan's value (creator estimate): https://www.youtube.com/watch?v=EIiXhCaZ4rw Claude on Sonnet 5.5: https://x.com/claudeai/status/2104633115620823187 Anthropic's Sonnet 5.5 page: https://www.anthropic.com/claude-sonnet-5-5 Artificial Analysis on Sonnet 5.5: https://x.com/ArtificialAnlys/status/2104640155843989864 Claude Devs, the Sonnet 5.5 guide: https://x.com/ClaudeDevs/status/2104687805876367793 the guide itself: https://claude.dev/blog/building-with-claude-sonnet-5-5/ Lenny's AI Builders - LAB 0010

    Wed, 30 Sep 2026
  • 53 - Claude Opus 5.5 vs GPT-6 Sol - which should builders use? | LAB 0006

    Opus 5.5 leads the Artificial Analysis chart. GPT-6 Sol has a lower published API list price. The half-price comparison refers to standard short-context API input and output list rates, not subscriptions or cost per finished task. But which one should you use for your own work? i look at the launches, the chart, cost per task, and one Sol miss in my own Notion planning. Then i show how to set up a five-task test for the jobs you actually do. i have not run Opus 5.5 or scored BuilderBench yet, so this episode does not pick a real-work winner. try this prompt: help me build a small benchmark for the work I actually do. Ask for my five recurring tasks. Turn each into one repeatable brief with a useful finished output, checks I can perform, and the time and cost to record. Mark what is still untested. Threadify, my own software, sponsors this episode. It helps you find potential leads in the replies to your Threads posts. See the plans and terms here: https://www.threadify.app/plans?utm_source=lenny-youtube&utm_medium=video&utm_campaign=lab-0006&utm_content=youtube-description-primary&video_slug=lab-0006&cta_slot=description&entry_angle=context&lp_variant=plans want to sponsor a future LAB episode? email sponsors@lennysaibuilders.com sources and tools: Anthropic on Claude Opus 5.5: https://www.anthropic.com/claude-opus-5-5 OpenAI's Sol/Luna announcement: https://openai.com/index/introducing-gpt-6-sol-and-luna/ GPT-6 Sol API list pricing: https://developers.openai.com/api/docs/models/gpt-6-sol Artificial Analysis index write-up: https://artificialanalysis.ai/articles/claude-opus-5-5/ Matt Pocock's Grill with Docs skill: https://github.com/mattpocock/skills/blob/main/skills/engineering/grill-with-docs/SKILL.md Lenny's AI Builders - LAB 0006 watch the episode: https://youtu.be/DAkQhc7Fh3U

    Mon, 28 Sep 2026
  • 52 - OpenAI says its ai solved over 100 maths problems. now what? | LAB 0005

    i looked at OpenAI's reported maths results, the GPT-6 Sol and Opus 5.5 price fight, Grok 4.7, Muse with Shopify and Qwen-Image-2.1. Then i gave Astra my existing LAB brand and asked for a matching thumbnail, banner and podcast cover. Astra directed the work in Codex; native ImageGen made the images. I show the finished set, the references and the checks. Give it your brand first. The maths result is OpenAI's claim, not an independently established result here. Some model details were reported or rumoured when this was recorded. Qwen's downloadable model needs suitable hardware. Threadify sponsors this episode. It helps turn replies on your Threads posts into lead signals for you to review: https://www.threadify.app/plans?utm_source=lenny-youtube&utm_medium=video&utm_campaign=lab-0005&utm_content=youtube-description-primary&video_slug=lab-0005&cta_slot=description&entry_angle=context&lp_variant=plans sources and updates: OpenAI updates: https://developers.openai.com/api/docs/changelog Claude Opus 5.5: https://www.anthropic.com/claude-opus-5-5 Fable plan limits: https://support.claude.com/en/articles/15424964-claude-fable-models-on-your-plan Muse announcement: https://about.fb.com/news/2026/09/introducing-muse-personal-ai-agent/ Qwen-Image-2.1 implementation: https://github.com/QwenLM/Qwen-Image-2.1 watch the episode: https://youtu.be/452rqqtYIyg

    Mon, 28 Sep 2026
Mostra altri episodi