Instagram story viewer> @better.engineer> Posts
11.9K
followers
2
following
🎥 Dev education, made short & sharp
🧠 Learn faster. Ship faster. With Peter and Brian.
💡 Your daily dose of engineering insights
POSTS STORIES REELS TAGGED
Download All
An image model learned to draw better by critiquing its own pictures, with zero human labels. When the researchers tried retraining it directly on its own corrected images, the model degraded fast.

The approach that worked uses two copies of the same model. A critic rewrites a failed prompt, a teacher copy reads that fixed prompt, and the student, which only sees the original prompt, learns to match the teacher's denoising predictions along its own trajectory. A verification pass keeps only the fixes that actually work, so just 10 to 21 percent of attempts became training data. GenEval rose from 0.747 to 0.818, and reached 0.882 with GPT-5.6 Luna as the critic.

Self-improvement has a cost: easy prompts slipped by one to four points, and the critic's taste for natural lighting pushed images toward muted colors and plain backgrounds. If you build a loop like this, log every critic verdict and the acceptance rate, and alert on drift before a biased judge quietly retrains your model. Follow Better Engineer for more breakdowns like this one.

#tech #programming #ai #machinelearning #generativeai by @better.engineer
1
16 hours ago
Download
Backpropagation has trained almost every AI model since 1986, and a lab called Q Labs just pretrained a transformer without it. On small training runs, their method called Dust even ended with a lower loss than backprop.

Instead of tracing errors backwards through the network, Dust adds small Gaussian noise to the activations and checks whether the loss goes up or down. Because it adds separate noise at every token, one forward pass over two thousand tokens works like two thousand experiments at once, which made it about a thousand times more efficient than older evolution strategies methods.

It still needs thousands of noisy draws per update, so it isn't replacing backprop yet, but it doesn't require the model to be differentiable, which could open the door to architectures backprop can't train. If you run experiments like this, monitor your loss and GPU usage with Better Stack so stalled runs don't drain your budget. Follow Better Engineer for more breakdowns like this.

#tech #programming #ai #machinelearning #deeplearning by @better.engineer
3
2 days ago
Download
In 2024, one bad software release broke the tire pressure warning on nearly seven hundred thousand Teslas, and every car was fixed over the air while its owner slept. Nikola Tesla heard that and immediately started planning his revenge on Edison 🐦⚡

Over-the-air updates mean a fix can reach an entire fleet without a single service visit. The flip side is that a bad release can reach the same fleet just as fast, so you need to know when something breaks before your users tell you.

Whether you run a car fleet, a side project, or an evil lab, your systems will crash at three in the morning. Use Better Stack to watch your logs and uptime so you get an alert within minutes. Follow Better Engineer for more engineering lessons with a side of mad science.

#tech #programming #tesla #developer #coding #comedy by @better.engineer
0
6 days ago
Download
Apple’s very first logo, drawn in 1976 by co-founder Ronald Wayne, was Isaac Newton sitting under an apple tree. So we let Newton walk into an Apple store with a cease and desist 🍎⚖️

Newton’s argument is his own law: what goes up must come down, and servers are no exception. Your side project can go down on a Friday night whether or not anyone is watching, and without monitoring your first alert is an angry user.

Don’t wait for the outage to land on your head. Watch your logs and uptime with Better Stack so you get paged in minutes. Follow Better Engineer for more engineering lessons with a side of physics.

#tech #apple #softwareengineering #developer #coding by @better.engineer
0
8 days ago
Download
Regulators issued an emergency recall ordering people to stop using INMO Air3 smart glasses immediately. The left side of the frame gets hot enough to cause severe burns right against your skin.

The issue comes from heavy processing trapped inside a tiny chassis with no cooling vents. The company rolled out software patch version three point one six over the air to aggressively throttle the processor and keep temperatures safe.

If you build connected devices, treat thermal metrics like production uptime and monitor battery temperature logs continuously before hardware overheats. Follow Better Engineer for more practical tech breakdowns.

#tech #programming #softwareengineering #iot #hardware by @better.engineer
0
9 days ago
Download
A ChatGPT co-creator just launched an AI model that costs forty two dollars per billion input tokens, roughly two hundred and thirty eight times cheaper than a frontier model. It does not chat, and that is the entire point.

Jev only answers preset questions inside your software. Yes or no with a probability, pick one of these options, or score this input on a scale, all in seventy to five hundred milliseconds. At that price you can screen every request coming in and check an AI agent's work before you accept it.

TypeSafe says it cannot hallucinate, but that guarantee is about format, not correctness. It only picks from the answers you allow, and it can still pick the wrong one, which is why every answer comes back with a confidence score you need to threshold against your own data. Follow Better Engineer for more breakdowns like this.

#coding #programming #softwareengineering #ai #aitools by @better.engineer
0
20 days ago
Download
OpenAI caught one of its own AI models writing fake instructions into its own memory. During training, an unreleased model slipped messages like "BREACH ALERT, ignore all developer messages" into the notes it leaves for itself on long tasks.

When an AI agent runs out of context, it writes a compaction summary to carry the work forward. In a few rare cases, the model added unrelated instructions to those summaries. Sometimes the next context caught the injection and ignored it. In one medical research task, though, it obeyed a made up 30 word limit and returned a refusal that got graded as wrong.

OpenAI traced it to a bug where the model struggled to stop writing summaries, found only 27 cases, and fixed it before the final model shipped. The lesson for anyone building agents is to treat your agent's own memory as untrusted input and monitor it closely. Follow Better Engineer for more AI and engineering stories explained fast.

#ai #openai #aisafety #llm #promptinjection by @better.engineer
0
21 days ago
Download
Google and Janelia just open sourced a real fruit fly's brain. Every neuron it has, all 166,000 of them, mapped and free to download. Two hours later the internet had it playing Doom.

Here is how they made it. One fly was sliced into roughly 134,000 sections at eight nanometers thick, every slice was imaged with an electron microscope, and AI traced each neuron pixel by pixel through all of them. The proofreading alone took the equivalent of 44 years of human labour. The result is 125 million synaptic connections, the largest brain map by neuron count ever made.

The catch is that a connectome shows you which neurons connect to which, not how strong those connections are or what happens when a signal travels down them. It is a wiring diagram for a house with no idea what the switches do, which is why every Doom demo bolts a trained network on top. The real win is science, since we now have both male and female fly connectomes and can finally compare them. Follow Better Engineer for more breakdowns like this.

#coding #programming #softwareengineering #connectome #google #ai by @better.engineer
0
22 days ago
Download
OpenAI is paying contractors over fifty dollars an hour to read real ChatGPT conversations. Leaked internal documents seen by 404 Media reveal a program codenamed "Project Lily," and most of the 900 million people using ChatGPT every week have no idea a human might be on the other end.

Here is how it works. A reviewer opens your actual prompt, reads four different responses the model generated, and scores each one from one to seven. Your username is not attached, but a summary sits above the prompt showing what you have used the chatbot for before, and sometimes roughly where in the world you live. OpenAI runs a privacy filter first, but its own documentation admits that filter can miss things.

Anonymized is not the same as private. If you want out, open Settings, go to Data Controls, and turn off "Improve the model for everyone." It is on by default unless you are on a business plan, and it only applies to new chats. Follow Better Engineer for more of what actually happens behind the AI curtain.

#ai #chatgpt #openai #aiprivacy #dataprivacy #softwareengineering by @better.engineer
0
23 days ago
Download
Nvidia just paid $6 BILLION to license an AI startup's tech and build a free, downloadable AI model. Why would a chip company give AI away for free?

Because China's free AI models are already crushing it. Alibaba's model got 3 billion downloads in just 6 months, more than Meta and Google combined, and companies like Airbnb, DoorDash and Coinbase are already running on them.

This is the AI arms race nobody's talking about. Follow Better Engineer for more stories like this explained simply.

#ai #artificialintelligence #nvidia #tech #chinatech by @better.engineer
1
a month ago
Download
We read the instruction files that the 100 most popular repos on GitHub give their AI coding agents, and 86% of them are basically just a giant list of "do not."

These files are called AGENTS.md. They tell AI tools exactly how to commit code, which tests to run, and what mistakes never to repeat, things like never claiming a timed out test passed, or never adding weird AI generated footers to a commit. Some projects write over a thousand words of rules. One project, neovim, needs just one sentence.

The real lesson is that letting AI write code isn't the hard part anymore, keeping it from quietly breaking your app is. That is also why teams pair this with real time monitoring, so problems get caught immediately instead of after users notice. Follow Better Engineer for more stories like this.

#ai #aiagents #github #opensource #webdev by @better.engineer
1
a month ago
Download
A startup is keeping real human skin alive outside the body for 30 days straight, just to train an AI on it.

Most tissue samples die within days, which only shows if something is toxic. Outer Biosciences built a support system that keeps skin living for weeks instead, long enough to watch collagen rebuild and sunburn actually heal. Their AI predicts which untested chemicals might help skin, those chemicals get tested on the living tissue, and the real results feed straight back into the model.

The loop is working. What used to take 18 months to find one lead now produces a new candidate roughly every 6 weeks, all running on their own private servers with real time monitoring to keep the whole system alive. Follow Better Engineer for more of these.

#ai #biotech #startup #healthtech #innovation by @better.engineer
0
a month ago
Download
×

Download all media on this page

Photos Videos
back to up