home.social

#simonwillison — Public Fediverse posts

Live and recent posts from across the Fediverse tagged #simonwillison, aggregated by home.social.

fetched live
  1. #SimonWillison discusses the impact of #AI on #softwareengineering. He highlights November 2025 as a turning point when #AIcodingagents became reliable. Willison also emphasises the need for #security measures against #promptinjection and predicts the rise of “dark factories” where AI autonomously generates and tests code. lennysnewsletter.com/p/an-ai-s #tech #media #news

  2. Simon Willison: A new way to extract detailed transcripts from Claude Code. “I’ve released claude-code-transcripts, a new Python CLI tool for converting Claude Code transcripts to detailed HTML pages that provide a better interface for understanding what Claude Code has done than even Claude Code itself. The resulting transcripts are also designed to be shared, using any static HTML hosting […]

    https://rbfirehose.com/2025/12/29/simon-willison-a-new-way-to-extract-detailed-transcripts-from-claude-code/
  3. Slop Acts of Kindness: Simon Willison on AI's "Helpful Hallucinations" – Why Verification Still Rules

    AI generates slop (fake but confident); kindness = polishing turds. Tools need human rails—GPT drafts great, fact-check kills flow [web:?? Simon link]
    Post-LLM truth: Latency thinking > token vomit.

    Devs: Cursor/Claude hacks? Prompt for citations always. Slop-proof your workflow.

    #AISlop #AIEthics #SimonWillison #PromptEngineering
    [SimonWillison.net - buff.ly/OWN812t ]

  4. 📝💡 Simon Willison bravely dives into the infinite #wisdom of GPT-3, discovering that the #AI can spit out text with all the #accuracy of a broken compass. His revelation? 🤔 You must fact-check everything it writes! Truly #groundbreaking stuff, Simon. 🌟
    simonwillison.net/2022/May/31/ #SimonWillison #GPT3 #FactCheck #HackerNews #ngated

  5. Meta’s surprise Llama 4 drop exposes the gap between AI ambition and reality - On Saturday, Meta released its newest Llama 4 multimodal AI models in a su... - arstechnica.com/ai/2025/04/met #machinelearning #simonwillison #biz#llama3 #llama4 #llama #meta #ai

  6. OpenAI threatens bans for probing new AI model’s “reasoning” process - Enlarge (credit: Andriy Onufriyenko via Getty Images)

    OpenAI t... - arstechnica.com/?p=2049959 #openaistrawberry #machinelearning #promptinjection #rileygoodside #simonwillison #aijailbreak #jailbreaks #o1-preview #strawberry #jailbreak #openaio1 #chatgpt #chatgtp #o1-mini #biz#gpt-4o #openai #gpt-3 #gpt-4 #hacks #ai #o1

  7. Before launching, GPT-4o broke records on chatbot leaderboard under a secret name - Enlarge (credit: Getty Images)

    On Monday, OpenAI employee Will... - arstechnica.com/?p=2024084 #largelanguagemodels #multimodalmodels #machinelearning #simonwillison #chatbotarena #gpt2-chatbot #gpt-4-turbo #aivibes #chatgpt #chatgtp #biz#gpt-4o #openai #gpt-4 #lmsys #ai

  8. Mysterious “gpt2-chatbot” AI model appears suddenly, confuses experts - Enlarge (credit: Getty Images)

    On Sunday, word began to spread... - arstechnica.com/?p=2020588 #machinelearning #simonwillison #aibenchmarks #chatbotarena #ethanmollick #gpt2-chatbot #samaltman #aivibes #gpt-3.5 #gpt-4.5 #biz#openai #gpt-3 #gpt-4 #gpt-5 #lmsys #ai

  9. Microsoft’s Phi-3 shows the surprising power of small, locally run AI language models - Enlarge (credit: Getty Images)

    On Tuesday, Microsoft announced... - arstechnica.com/?p=2019278 #largelanguagemodels #smalllanguagemodels #machinelearning #simonwillison #microsoftphi #microsoft #chatgpt #chatgtp #mistral #biz#openai #phi-2 #phi-3 #ai