#geekstuff — Public Fediverse posts
Live and recent posts from across the Fediverse tagged #geekstuff, aggregated by home.social.
-
Software Engineering fundamentals matter more than ever
The manifestation of my imposter syndrome, for me and today, is what does it mean to be a software engineer. There’s a lot more noise than signal on the Internet about agentic engineering, what can be accomplished, and its implications for the future. The title I chose rather gives it away; it’s about choosing — carefully — all the things you need to choose when you’re solving the puzzles of software and systems development.
Beyond the hype and junkie-like marketing fervor of “major model providers”, I found a really interesting power tool with the combination of harness and models. I’ve been following how friends have been using these tools, and learning a ton. As usual, the folks doing some of the most amazing things aren’t the ones crowing about it, or posting narrative blurbs in social media about the end of this profession. They found a “big damn stick”, they’re exploring the fulcrum points, and they’re representing good ole Archimedes to lean into that lever, moving the world.
In the past year, agent harnesses crossed the “can it be done” rubicon. (yep, jumping forward to Roman references). I would not have wished for the world’s knowledge to taken without permission and regard, or the lunatics to delve into economic self-dealing that’s peanut buttering over the otherwise tanking US economy. The economic models for the large models aren’t viable from any report that I’ve seen, but the capability isn’t going away. Instead it’s shrinking (fast!). Open weight models are making (beefy) personal computers quite capable of doing the same. They’re not quite as effective, but the delta in time and capability isn’t large.
“Can it be done” is only the start, not even close to the majority a software or system engineer’s profession. It’s like when I learned to weld in my 20’s – I quickly created things that I couldn’t lift or even get out the door of the shop. (thank goodness for acetylene torches). What I learned then is I think the same lesson, different medium: How something goes together is what makes all the difference.
If you use agentic harnesses to develop with a bit of foresight, you can get not only “it works”, but also “it’s testable” (I heavily lean into the prompt “develop with red/green TDD”). But it’s not very solid much above that. The seams — how your code works, it’s “API”, and how it fits with other software — are as much art as science. It is made up of subjective measures that rely on your viewpoint (and experience, as well as your guesses) for both what you’re solving now, and how to live with that software over a long period of time.
Making software debuggable, maintainable, layered, and composable – that’s still quite a trick. Quite a lot of that work requires extensive, thoughtful reasoning. And that’s where the LLM’s today, even the leading edge of the “capability” from frontier models, fall short.
It helps to know that LLMs don’t “reason”. They predict, and the models themselves are effectively written human knowledge compressed. So if it’s in human knowledge that was encoded, it can echo out the human reasoning. For agents focused on software development, those reasoning traces are the precious data for the models. There’s a very approachable research paper on just how bad LLMS are at reasoning called The Illusion of Thinking. There is some research I’m following that includes prediction of results of actions, but that’s not what we have today with coding agents. It’s a pretty different – and fascinating – area of research. If you want to explore, go digging on how “JEPA models” work, LeWorld Model, and recent talks by Yann LeCun.
While you’re working with LLMs though, there’s still a ton of ways to make them more effective. I think there’s a lot of advances that we haven’t even really begun to eek out. Most of the wins I’m seeing today involve providing it good, concise data to work from, at the right time, and providing deterministic validation tooling with natural language feedback that the LLM can use to correct itself. The amazing thing to me isn’t that it can predict what to write, but that it is effective at tool calling and following instructions.
Another downside of this instruction following is what Simon Willison coined as the lethal trifecta. Basically – LLM models can’t distinguish between good advice and bad. They’re foundationally incapable of always and consistently preventing prompt injection attacks. “Alignment work”, safety harnesses, and sandboxes all help to add barriers against the worst, but there are fundamental gaps. And frankly, something that tirelessly follows instructions without having good reasoning is nightmare fuel to me.
I hope there will be near-term nadvances in how models are trained to include the equivalent of reasoning traces for post-training (RLHF). In my ideal future, these include more of what it means to build software with clean interfaces, that’s debuggable, and and that’s maintainable as a key part of the reinforced evaluations. Carefully reviewing, planning, and fixing the seams of software (and systems) is one of the critical skills we both can, and need to, employ when developing software – with or without agentic assistants. And as I see the wave of “Oh, that’s easy to implement…” and people reaching for clankers to get it done, I think it’s more important than ever.
It’s a great time to be following folks who write, talk, and share about the craft of software, and how we can be better artisans. Hopefully it’s obvious, but there’s never a single answer — a panacea. It’s always about tradeoffs, choosing what makes sense for the problem at hand. With the help of a lot of great minds sharing their thoughts — both now and going back decades — we have a great tool chest for this work. It’s about picking, or reworking to move to a better choice, the right abstractions. It’s core is managing the cognitive load, learning which pieces we need to be stable, and where we want our work to flex and bend (and how).
#AI #Geekstuff #LLM #MLAnd yes, I wrote the damn em-dashes myself. I’m too in love with a recursive parenthetical in my writing, and I like a break from commas and parentheses.
-
I challenge everyone - what is the oldest (software) raid that you are still running? Send a proof of metadata. I've a nice one that I am going to reveal in a few days. #challenge #forfun #geekstuff #uptime
-
C’est open-source, c’est efficace, et ça laisse plus de temps pour boire un café ☕️
#BlablaLinux #Python #WikiJS #Automation #Linux #GeekStuff #SysAdmin
-
neues Video:
#LILYGO T-Echo #Meshtastic mit #BME280
Vorstellung des Gerätes und erster Erfahrungsbericht, Konfiguration über Bluetooth und Menü:
--> https://youtu.be/eiC732NR8Kw
#TECHO #LoRa #GeekStuff #Maker #Funk #868MHz #Messenger #Funknetzwerk
-
I just optimised some code:
https://www.dwitter.net/d/33418
This is a #realtime #pixelshader in #javascript - a fun piece of #graphics #compsci and #geekstuff
-
Toujours impressionné par la qualité du travail de BunsenLabs Linux ( @bunsenlabs ) ! 👏
J'ai mis à jour un eee PC 900 (i386) de 2009 vers leur dernière version Boron
👇
https://www.bunsenlabs.org/Basée sur la puissante & mighty Debian :debian: 🤘
Doté de Kernel pae 6.1 ! :apartyblobcat:
Tout fonctionne parfaitement, même sur cette vieille micro-machine riquiqui ! 👍
ça en fait un terminal minimaliste très geek, avec une interface modulable et un système Openbox rapide et ergonomique.
#BunsenLabs #Linux #Debian #Openbox #eeePC #vintage #GeekStuff
-
Our recent home renovation work had me concerned about air quality. So I picked up this Tuya air quality meter. It has a nice assortment of sensors measuring particulates at 10μm, 2.5μm, 1μm, and 0.3μm. It also measures CO₂, CO, HCHO (formaldehyde), VOCs, and shows the EPA Air Quality Index (ACI). I'm using it stand-alone but it can pair with an Android app or be used as a sensor via WiFi with an open API
.
#photooftheday #airquality #airmeter #geekstuff -
The new monitor stand for my old Acer 27" display finally arrived. I assembled it before dawn this morning, and now have it in place for my Linux box.
It's like having a brand new display on the cheap! I get weirdly excited about giving older hardware new life. #GeekStuff
-
Welp, I just ordered a new-to-me Mac mini.
I always order refurbished Macs to save tons of dough, so I'm replacing my late 2012 model with a late 2018 one that's maxed out with 64GB of RAM, and cuz it has very little internal storage space, I'm also getting a 6-terabyte (!) external mini-sized drive. All of this should keep me going for most of the next decade.
My 2012 model just drags so much. I'll continue to use it for some lightweight tasks, but #musicproduction is getting a new home.
-
I'm upgrading our network storage at home with a new 18TB RAID box this weekend
.
#photooftheday #nas #geekstuff -
Could Kelly Rowland have used the =HYPERLINK() function to message Nelly?
The Kelly Rowland/Nelly song Dilemma features an infamous scene amongst nerds where Kelly Rowland tries to send a message to Nelly using a Nokia 9210 Communicator.Unfortunately, she does this using the built in spreadsheet program and receives no reply.
People suggested she might be using
-
The most useless clock I ever made.... (It only displays time in obscure formats, and in the serial terminal).
UTC (Ok not obscure & probably the most useful on the list), New Earth Time, Decimal (2 different versions), Dalek Rel Time, French Fractional Time, Beats Time, Hexadecimal Clock, Star Trek Stardate#Time #Clock #Useless #ESP32 #UTC #Beats #FractionalTime #DecimalTime #NET #nerdStuff #geekStuff #Stardate
-
I started the day following a bit above 1300 people.
Removed a lot of accounts that haven't updated in 5-10 years.
Removed corporate accounts, and a dumptruck full of Toronto politics, US politicians, and other garbage.
Mostly left with musicians, artists, film makers. #geekstuff
I expect I haven't really scratched the surface.
Reminded that Peter Davison was chased off Twitter in 2017 over suggesting that maybe Jodie Whittaker might be a good change for the Dr. Who franchise.
-
CW: spoiler for ep 10
Picard's couple of days with Jack would be a sufficient connection to use.
-Yes, Shaw's final evaluation was heart-warming, but I didn't think quite enough groundwork had been laid for it in the earlier part of the season. It was a crucial bit too radical a change from everything else we'd seen of him.
-
OpenAI announces GPT-4
We’ve created GPT-4, the latest milestone in OpenAI’s effort in scaling up deep learning. GPT-4 is a large multimodal model (accepting image and text inputs, emitting text outputs) that, while less capable than humans in many real-world scenarios, exhibits human-level performance on various professional and academic benchmarks. For example, it passes a simulated bar exam with a score arou
-
Sure, I’ll play along.
To help make #Mastodon connections: list 5-7 things that aren't in your profile but that interest you as #tags so they are searchable. Then boost this post or repeat its instructions so others know to do the same. 👍#psytrance
#Filk
#Cooking
#geekstuff
#musicaltheatre
#canadianpolitics
#polyamory
#tattoos -
Some maitenance need in the house then onto frankensteining my old rig back to life with some watercooling !
#geekstuff always brings a smile on my face 😁 -
Finally learning #tmux and I seriously wonder why I didn't do that earlier. Beats #screen in a lot of ways. #geekstuff