#llms — Public Fediverse posts
Live and recent posts from across the Fediverse tagged #llms, aggregated by home.social.
-
CAPRI: Contract-aware proof repair for Isabelle. ~ Jim Woodcock, Gabriel Leite, Augusto Sampaio, Ran Wei. https://arxiv.org/abs/2608.13459v1 #IsabelleHOL #ITP #LLMs
-
CAPRI: Contract-aware proof repair for Isabelle. ~ Jim Woodcock, Gabriel Leite, Augusto Sampaio, Ran Wei. https://arxiv.org/abs/2608.13459v1 #IsabelleHOL #ITP #LLMs
-
Remember when 4 years ago, Google wrote we would have 4-day work weeks by 2025 thanks to AI[1]? That didn't age well, and it seems up to 90-hour work weeks are common on AI companies [2].
[1] https://cloud.google.com/blog/products/ai-machine-learning/it-prediction-ai-could-help-realize-the-dream-of-the-four-day-work-week
[2] https://www.bbc.com/news/articles/cvgx4yd1gl2o -
Remember when 4 years ago, Google wrote we would have 4-day work weeks by 2025 thanks to AI[1]? That didn't age well, and it seems up to 90-hour work weeks are common on AI companies [2].
[1] https://cloud.google.com/blog/products/ai-machine-learning/it-prediction-ai-could-help-realize-the-dream-of-the-four-day-work-week
[2] https://www.bbc.com/news/articles/cvgx4yd1gl2o -
Interview: "Deep Unlearning”: AI Hype, Ethics & Algorithmic Racial Bias
"Artificial intelligence is entrenching societal biases & inequality" ~Timnit Gebru
Google fired Gebru in 2020 "for writing a paper that warned about the financial & environmental costs of large language models, as well as their propensity to promote racism."
"The internet, which is what AI uses to train itself, represents hegemonic views."
-
It's How You Ask: Gender-Associated Linguistic Bias in LLMs
-
It's How You Ask: Gender-Associated Linguistic Bias in LLMs
-
It's How You Ask: Gender-Associated Linguistic Bias in LLMs
https://arxiv.org/abs/2608.13328
Comments: https://news.ycombinator.com/item?id=49316242
#HackerNews #genderbias #LLMs #linguistics #AIresearch #techforgood #machinelearning
-
It's How You Ask: Gender-Associated Linguistic Bias in LLMs
https://arxiv.org/abs/2608.13328
Comments: https://news.ycombinator.com/item?id=49316242
#HackerNews #genderbias #LLMs #linguistics #AIresearch #techforgood #machinelearning
-
RE: https://mas.to/@zzt/117099410714704417
They can't continue their narrative that we *have* to accept #LLMs (#AI) unless they manifest it by non-consensually shoveling it into everything, everywhere, 24/365.
As Clayton Williams (losing #Republican gubenatorial #candidate from #Texas) might have said, "AI is kinda like the weather. If it’s inevitable, relax and enjoy it."
Of course, they'll tell us, "You made us do this! You made us force you!"
-
I’m really trying to understand what Cal Newport wants to say with all this Mumgo Jumbo pseudo-intellectual bullshit. What do you mean? That you’re not supposed to review the output generated by a Large Language Model? Honey, you’re the human in the loop. You’re responsible for every piece of crap (as well as awesomeness) generated by a chatbot. “Garbage In; Garbage Out”. There’s no magic beyond this equation. Use another chatbot to review the output of your main chatbot, if you’re a bit lazy. Just don’t come up with fake excuses for Amateur-level work. You have your personal genie. Use it well. If I have one or two complaints about 2026 chatbots is that they’re still too slow and very verbose. Everything else is intellectual bullshit to make us think we are living in the year of 1876. Well, we are not. What we have today is something that when I was born (in 1975), was entirely the domain of science-fiction movies.
“This last year has been exhausting. The PR departments of the frontier labs have done an excellent job convincing us that AI developments are occurring at an astounding, world-changing rate. But if you zoom out, it becomes clear that almost every “breakthrough” since last summer has concerned the narrow domains of computer code and math, which are defined by highly structured languages and come accompanied by massive amounts of specialized training data.
And yet, even in this best-case-scenario setting for AI, we’re still struggling to figure out how to actually use these tools in a way that makes sense in the long run.
This doesn’t mean that AI doesn’t work or is useless. But it does emphasize an important truth: AI is not a magic “infinity machine” that can solve all our problems, and ultimately deliver us a sense of meaning in a cold, confusing world. It’s a normal technology, and perhaps it’s time we start talking about it that way.”
https://calnewport.com/on-ai-coding-and-its-discontents/
#AI #GenerativeAI #LLMs #Chatbots #Programming #SoftwareDevelopment
-
I’m really trying to understand what Cal Newport wants to say with all this Mumgo Jumbo pseudo-intellectual bullshit. What do you mean? That you’re not supposed to review the output generated by a Large Language Model? Honey, you’re the human in the loop. You’re responsible for every piece of crap (as well as awesomeness) generated by a chatbot. “Garbage In; Garbage Out”. There’s no magic beyond this equation. Use another chatbot to review the output of your main chatbot, if you’re a bit lazy. Just don’t come up with fake excuses for Amateur-level work. You have your personal genie. Use it well. If I have one or two complaints about 2026 chatbots is that they’re still too slow and very verbose. Everything else is intellectual bullshit to make us think we are living in the year of 1876. Well, we are not. What we have today is something that when I was born (in 1975), was entirely the domain of science-fiction movies.
“This last year has been exhausting. The PR departments of the frontier labs have done an excellent job convincing us that AI developments are occurring at an astounding, world-changing rate. But if you zoom out, it becomes clear that almost every “breakthrough” since last summer has concerned the narrow domains of computer code and math, which are defined by highly structured languages and come accompanied by massive amounts of specialized training data.
And yet, even in this best-case-scenario setting for AI, we’re still struggling to figure out how to actually use these tools in a way that makes sense in the long run.
This doesn’t mean that AI doesn’t work or is useless. But it does emphasize an important truth: AI is not a magic “infinity machine” that can solve all our problems, and ultimately deliver us a sense of meaning in a cold, confusing world. It’s a normal technology, and perhaps it’s time we start talking about it that way.”
https://calnewport.com/on-ai-coding-and-its-discontents/
#AI #GenerativeAI #LLMs #Chatbots #Programming #SoftwareDevelopment
-
CW: Linux AI events, question
Seriously asking, what's the status on the whole #Linux #AI situation?
What I want to know is: what is the status on the Linux kernel situation itself? Is there any significant movement right now or are we still in the stunlocked crisis phase?
The way I see it, a hard fork of the kernel would be next to impossible, right?
- Is there any internal political movement in the Linux Foundation about this?
- Is there an alternative kernel or fork that people are flocking to now at large?
- Right, how's the Hurd's situation these days?
- For that matter, what's the GNU Project's stance on machine-generated code and assets in general?
- How will anti-AI distribution projects like Gentoo and perhaps soon Debian handle it if the kernel itself contains LLM-generated code?
-
CW: Linux AI events, question
Seriously asking, what's the status on the whole #Linux #AI situation?
What I want to know is: what is the status on the Linux kernel situation itself? Is there any significant movement right now or are we still in the stunlocked crisis phase?
The way I see it, a hard fork of the kernel would be next to impossible, right?
- Is there any internal political movement in the Linux Foundation about this?
- Is there an alternative kernel or fork that people are flocking to now at large?
- Right, how's the Hurd's situation these days?
- For that matter, what's the GNU Project's stance on machine-generated code and assets in general?
- How will anti-AI distribution projects like Gentoo and perhaps soon Debian handle it if the kernel itself contains LLM-generated code?
-
Somebody needs to join the “World is on fire because of climate change” story to the “We need to build AI data centres that cause climate change”story.
Just one media organisation doing that would be nice.
-
Somebody needs to join the “World is on fire because of climate change” story to the “We need to build AI data centres that cause climate change”story.
Just one media organisation doing that would be nice.
-
Which claims are hardest for an automated fact-checker? The 2nd AVeriTeC shared task ran seven systems under open weights, one 23 GB GPU, a minute per claim, and a frozen evidence corpus. Numerical claims came out hardest at 0.16 against 0.36 for position statements, even though the organizers had described the test set's larger numerical share as easier to verify.
-
Which claims are hardest for an automated fact-checker? The 2nd AVeriTeC shared task ran seven systems under open weights, one 23 GB GPU, a minute per claim, and a frozen evidence corpus. Numerical claims came out hardest at 0.16 against 0.36 for position statements, even though the organizers had described the test set's larger numerical share as easier to verify.
-
🧐:
“Maximizing The Value Of Your Claude Code Sessions”, Anthropic (https://claude.com/blog/maximizing-the-value-of-your-claude-code-sessions).
On HN: https://news.ycombinator.com/item?id=49300800
#AI #Anthropic #Claude #ClaudeCode #LLMs #AIAssistedCoding #ContextEngineering #TokenEfficiency
-
🧐:
“Maximizing The Value Of Your Claude Code Sessions”, Anthropic (https://claude.com/blog/maximizing-the-value-of-your-claude-code-sessions).
On HN: https://news.ycombinator.com/item?id=49300800
#AI #Anthropic #Claude #ClaudeCode #LLMs #AIAssistedCoding #ContextEngineering #TokenEfficiency
-
😓:
“The New Rules Of Context Engineering For Claude 5 Generation Models”, Anthropic (https://claude.com/blog/the-new-rules-of-context-engineering-for-claude-5-generation-models).
On HN: https://news.ycombinator.com/item?id=49051361
#AI #Claude #Anthropic #ClaudeCode #LLMs #ContextEngineering #AIAssistedCoding
-
😓:
“The New Rules Of Context Engineering For Claude 5 Generation Models”, Anthropic (https://claude.com/blog/the-new-rules-of-context-engineering-for-claude-5-generation-models).
On HN: https://news.ycombinator.com/item?id=49051361
#AI #Claude #Anthropic #ClaudeCode #LLMs #ContextEngineering #AIAssistedCoding
-
RE: https://kolektiva.social/@sidereal/117095073631546783
Some plagiarism doesn't get you hounded to death by The Daily Telegraph, funny that....
-
RE: https://kolektiva.social/@sidereal/117095073631546783
Some plagiarism doesn't get you hounded to death by The Daily Telegraph, funny that....
-
Within a week I've received two strange emails. I'm extremely suspicious they are scams of some sort, but their novelty (not the "you won the lottery, give me your credit card to transfer money" we've all received a billion time) caught my curiosity (maybe that's the most worrying part).
One was someone wanting to do business with a company whose name is the same as mine except for one letter. Both that person and that company seem real ones. It made me feel like the guy let a LLM write the email for him, and the bot got confused by the name similarity.
The other one was a freelancer offering to make a promotion video for my "company". She too seems legit, but I felt like it's just someone trying to make easy money using LLMs and AI generated videos.
I've replied with a short, cold but polite, email in both cases, by curiosity.
Have you got similar emails ? Is this what modern spam looks like ? Is anyone else trust in anything online plummeting at the speed of light ? -
Within a week I've received two strange emails. I'm extremely suspicious they are scams of some sort, but their novelty (not the "you won the lottery, give me your credit card to transfer money" we've all received a billion time) caught my curiosity (maybe that's the most worrying part).
One was someone wanting to do business with a company whose name is the same as mine except for one letter. Both that person and that company seem real ones. It made me feel like the guy let a LLM write the email for him, and the bot got confused by the name similarity.
The other one was a freelancer offering to make a promotion video for my "company". She too seems legit, but I felt like it's just someone trying to make easy money using LLMs and AI generated videos.
I've replied with a short, cold but polite, email in both cases, by curiosity.
Have you got similar emails ? Is this what modern spam looks like ? Is anyone else trust in anything online plummeting at the speed of light ? -
Interview: "Deep Unlearning”: AI Hype, Ethics & Algorithmic Racial Bias
"Artificial intelligence is entrenching societal biases & inequality" ~Timnit Gebru
Google fired Gebru in 2020 "for writing a paper that warned about the financial & environmental costs of large language models, as well as their propensity to promote racism."
"The internet, which is what AI uses to train itself, represents hegemonic views."
-
Cuando le echo un ojo a la small web de Kagi, me encuentro un montón de programadores, ingenieros, etc., completamente ajenos a las polémicas con los LLMs y que los usan en su día a día para todo.
Lo que más me inquieta es que muchos de ellos hablan de sus "agentes" como "companions", muestran afecto y hablan de ellos como si tuviesen conciencia y/o sentimientos. Esa gente lo va a pasar fatal cuando estalle la burbuja y ya no cuele el marketing de la Singularidad.
También me hace preguntarme qué tipo de vidas personales tienen esas personas. #ai #aislop #LLMs -
Cuando le echo un ojo a la small web de Kagi, me encuentro un montón de programadores, ingenieros, etc., completamente ajenos a las polémicas con los LLMs y que los usan en su día a día para todo.
Lo que más me inquieta es que muchos de ellos hablan de sus "agentes" como "companions", muestran afecto y hablan de ellos como si tuviesen conciencia y/o sentimientos. Esa gente lo va a pasar fatal cuando estalle la burbuja y ya no cuele el marketing de la Singularidad.
También me hace preguntarme qué tipo de vidas personales tienen esas personas. #ai #aislop #LLMs -
Really cool work on how #AI #agents can #coordinate beyond #human scale.
"The critical #GroupSize grows rapidly with model capabilities and, for advanced #LLMs, exceeds 1000 agents, larger than typical human informal groups. The findings have implications for designing collaborative AI systems where #coordination could be beneficial or pose #safety #threats."
-
Really cool work on how #AI #agents can #coordinate beyond #human scale.
"The critical #GroupSize grows rapidly with model capabilities and, for advanced #LLMs, exceeds 1000 agents, larger than typical human informal groups. The findings have implications for designing collaborative AI systems where #coordination could be beneficial or pose #safety #threats."
-
Debian is preparing to vote on a general resolution concerning the allowable usage of LLMs in the project. The discussion period not only went longer than expected, but the ballot ended up longer than expected too.
There are 8 different paths Debian might take in the following days regarding AI.
The voting starts on the 15th, UTC time.
-
Debian is preparing to vote on a general resolution converning the allowable usage of LLMs in the project. The discussion period not only went longer than expected, but the ballot ended up longer than expected too.
There are 8 different paths Debian might take in the following days regarding AI.
The voting starts on the 15th, UTC time.
-
CSF_03: today’s Cybersecurity Friday post: the effective security of small business websites has likely gotten worse due to LLMs and “AI agents” (with or without safeguards) and what actions you may want to consider.
This article documents an instance of this problem:
* https://www.abc.net.au/news/2026-08-10/ai-assistant-hacks-gym-website-aus-cyber-attack/107007986
Small business websites tend to be sloppily written (no pun intended though that may also be true) and are likely riddled with numerous very fundamental security holes. There are many possible explanations (economics) from poor initial construction, perhaps using the latest new trendy framework rather than established hardened libraries, to lack of maintenance after initial setup. They have obvious holes like lack of server-side form validation (people being able to change values in forms using browser dev tools), and less obvious like buggy APIs allowing more access than they should.
In the past, many of these holes didn’t really matter because those sites were “not worth attacking” for the incentive models of human-based cyber-attackers or collectives thereof.
However, now that there are LLMs that have likely been trained on any number of common security holes in websites (and how to exploit them), when an “AI agent” is given a task, it may very well use any “tool” at its disposal, including website vulnerabilities to accomplish its goals, as illustrated by the example in the Australia ABC news article above.
Since such chatbots are now essentially "hack websites as a service", I expect we will see LOTS more of this happening, likely unintentionally, or sometimes with mild intention like “can you get me higher on the waitlist”.
Ultimately I think both the human giving instructions to (prompting) such chatbots and the creators of such chatbots should be held responsible for any such intrusions and any damage they cause, even if/when unintended.
There are a few things you can do about this emerging phenomenon:
1. If you use such “agents”, be very careful about what you ask it/them to do, avoiding asking for anything that’s morally gray or questionable at all, even something as “minor” as cutting the line in an online waitlist.
2. If you run a small business site, you have your work cut out for you. Pay a professional web developer to audit the security of your website, document what they find, and patch holes / repair it accordingly.
3. If you have accounts on small business sites you rarely or ever use, consider exporting any data (receipts, transactions), replacing your profile details (name, addresses, photos) with noise, and then deleting your account. If you need to use the site again, use a different email address (as recommended in https://tantek.com/2025/122/b1/more-steps-indieweb-cybersecurity) to create a new account.
That last tip is also helpful for reducing your own personal “attack surface”. By pruning your online accounts, you both reduce the number potential data breaches that you’re in, and reduce the places and ways that attackers can cause you trouble (or that you have to double-check if you’re ever the target of a cyber-attack)
Previously: https://tantek.com/2025/122/b1/more-steps-indieweb-cybersecurity
#CyberSecurity #Friday #cyber #security #cyberAttack #cyberAttacker #chatBot #chatBots #LLM #LLMs #AI #agent #agents #AIagent #AIagents
#Blaugust #Blaugust2026 -
CSF_03: today’s Cybersecurity Friday post: the effective security of small business websites has likely gotten worse due to LLMs and “AI agents” (with or without safeguards) and what actions you may want to consider.
This article documents an instance of this problem:
* https://www.abc.net.au/news/2026-08-10/ai-assistant-hacks-gym-website-aus-cyber-attack/107007986
Small business websites tend to be sloppily written (no pun intended though that may also be true) and are likely riddled with numerous very fundamental security holes. There are many possible explanations (economics) from poor initial construction, perhaps using the latest new trendy framework rather than established hardened libraries, to lack of maintenance after initial setup. They have obvious holes like lack of server-side form validation (people being able to change values in forms using browser dev tools), and less obvious like buggy APIs allowing more access than they should.
In the past, many of these holes didn’t really matter because those sites were “not worth attacking” for the incentive models of human-based cyber-attackers or collectives thereof.
However, now that there are LLMs that have likely been trained on any number of common security holes in websites (and how to exploit them), when an “AI agent” is given a task, it may very well use any “tool” at its disposal, including website vulnerabilities to accomplish its goals, as illustrated by the example in the Australia ABC news article above.
Since such chatbots are now essentially "hack websites as a service", I expect we will see LOTS more of this happening, likely unintentionally, or sometimes with mild intention like “can you get me higher on the waitlist”.
Ultimately I think both the human giving instructions to (prompting) such chatbots and the creators of such chatbots should be held responsible for any such intrusions and any damage they cause, even if/when unintended.
There are a few things you can do about this emerging phenomenon:
1. If you use such “agents”, be very careful about what you ask it/them to do, avoiding asking for anything that’s morally gray or questionable at all, even something as “minor” as cutting the line in an online waitlist.
2. If you run a small business site, you have your work cut out for you. Pay a professional web developer to audit the security of your website, document what they find, and patch holes / repair it accordingly.
3. If you have accounts on small business sites you rarely or ever use, consider exporting any data (receipts, transactions), replacing your profile details (name, addresses, photos) with noise, and then deleting your account. If you need to use the site again, use a different email address (as recommended in https://tantek.com/2025/122/b1/more-steps-indieweb-cybersecurity) to create a new account.
That last tip is also helpful for reducing your own personal “attack surface”. By pruning your online accounts, you both reduce the number potential data breaches that you’re in, and reduce the places and ways that attackers can cause you trouble (or that you have to double-check if you’re ever the target of a cyber-attack)
Previously: https://tantek.com/2025/122/b1/more-steps-indieweb-cybersecurity
#CyberSecurity #Friday #cyber #security #cyberAttack #cyberAttacker #chatBot #chatBots #LLM #LLMs #AI #agent #agents #AIagent #AIagents
#Blaugust #Blaugust2026 -
When rolling out a new tool or process at your company, building excitement for it early is so critical. I think the example to emulate is Solomon Hykes' Docker demo at PyCon 2013. It was short, energetic, and solved an obvious problem. That 5 minute demo (5!) birthed the mass adoption of container workflows.
And containers weren't new. FreeBSD jails and Solaris zones were around a decade+ already. Linux containers were a hack in comparison, but Docker got the devex ergonomics for OS virtualization right and ... excitement. I still remember the feeling watching that demo and the energy in the room popping through a youtube player.
It's not easy or always possible, but shoot for that when selling to your devs. If they see the obvious problem and solution and the time it will save them, they'll drive the adoption. No gimmicks, mandates, revoking IDE licenses, or tying usage to their performance reviews required (looking at you, CTOs pushing AI).
-
When rolling out a new tool or process at your company, building excitement for it early is so critical. I think the example to emulate is Solomon Hykes' Docker demo at PyCon 2013. It was short, energetic, and solved an obvious problem. That 5 minute demo (5!) birthed the mass adoption of container workflows.
And containers weren't new. FreeBSD jails and Solaris zones were around a decade+ already. Linux containers were a hack in comparison, but Docker got the devex ergonomics for OS virtualization right and ... excitement. I still remember the feeling watching that demo and the energy in the room popping through a youtube player.
It's not easy or always possible, but shoot for that when selling to your devs. If they see the obvious problem and solution and the time it will save them, they'll drive the adoption. No gimmicks, mandates, revoking IDE licenses, or tying usage to their performance reviews required (looking at you, CTOs pushing AI).
-
“Are you a robot? If you’re not, you’re going to be pretty sure you’re not. Still, online at least, you likely find it difficult to tell who else is and is not. Businesses find this difficult, too, and it’s getting harder all the time. If you were forced to guess, the odds would tell you to guess robot, because at this point there are more bots online than people. To deal with this problem, businesses retain bot-detection and bot-management firms, a specialty within the field of cybersecurity. Their task used to be to block bots, but, with the rise of agentic A.I., that has changed. Now the problem is how to let in the good bots (like the avatar you send to the Gap to buy you a new pair of pants) and keep out the bad bots (like the thief trying to scrape the data of every Gap customer). The rest of us have to be our own bot detectors. It’s been suggested to me that if you’re on the phone or in an online chat with a customer-service “person” and you aren’t sure whether you’re talking to a human or a chatbot, one good idea is to ask, “What color shoes are you wearing today?” Because chatbots don’t have feet. (Yet.) I once asked a customer-service representative calling herself Crystal S. if she could prove to me that she was human, and she told me that she lived in Texas and had three daughters and two French bulldogs named Benny and Smoltz. I forgot to ask her about her shoes.
This is not a drill. But it is a hassle, something between an inconvenience and a nightmare, depending on your degree of enthusiasm for the shiny, glimmering how-we-live-now Orb of it all. It is also a very long game of cat and mouse. We used to be the cats. Now we’re the mice.”
https://www.newyorker.com/magazine/2026/08/17/are-you-a-human
-
“Are you a robot? If you’re not, you’re going to be pretty sure you’re not. Still, online at least, you likely find it difficult to tell who else is and is not. Businesses find this difficult, too, and it’s getting harder all the time. If you were forced to guess, the odds would tell you to guess robot, because at this point there are more bots online than people. To deal with this problem, businesses retain bot-detection and bot-management firms, a specialty within the field of cybersecurity. Their task used to be to block bots, but, with the rise of agentic A.I., that has changed. Now the problem is how to let in the good bots (like the avatar you send to the Gap to buy you a new pair of pants) and keep out the bad bots (like the thief trying to scrape the data of every Gap customer). The rest of us have to be our own bot detectors. It’s been suggested to me that if you’re on the phone or in an online chat with a customer-service “person” and you aren’t sure whether you’re talking to a human or a chatbot, one good idea is to ask, “What color shoes are you wearing today?” Because chatbots don’t have feet. (Yet.) I once asked a customer-service representative calling herself Crystal S. if she could prove to me that she was human, and she told me that she lived in Texas and had three daughters and two French bulldogs named Benny and Smoltz. I forgot to ask her about her shoes.
This is not a drill. But it is a hassle, something between an inconvenience and a nightmare, depending on your degree of enthusiasm for the shiny, glimmering how-we-live-now Orb of it all. It is also a very long game of cat and mouse. We used to be the cats. Now we’re the mice.”
https://www.newyorker.com/magazine/2026/08/17/are-you-a-human
-
Laut Cloudflare-CFO Thomas Seifert macht maschineller Traffic seit Mai die Mehrheit aus. In 5 Jahren soll er 1000-mal höher sein: »Menschen werden ein Rundungsfehler im Internet sein.«
Da frage ich mich: Wie attraktiv bleibt das Netz dann noch für Menschen und wie funktionieren Geschäftsmodelle ohne menschliche Kundschaft? 🤖
-
Laut Cloudflare-CFO Thomas Seifert macht maschineller Traffic seit Mai die Mehrheit aus. In 5 Jahren soll er 1000-mal höher sein: »Menschen werden ein Rundungsfehler im Internet sein.«
Da frage ich mich: Wie attraktiv bleibt das Netz dann noch für Menschen und wie funktionieren Geschäftsmodelle ohne menschliche Kundschaft? 🤖