ByteByteGo Newsletter
Subscribe
Sign in
Home
Sponsoring ByteByteGo
Become an AI Engineer Cohort
Archive
About
Latest
Top
Discussions
How to Run a Big Model on Cheap Hardware?
A large AI model can run on modest hardware only by reducing the memory it occupies, reducing the calculations it performs, or moving some work to…
13 hrs ago
•
ByteByteGo
237
3
9
EP226: API Concepts Every Software Engineer Should Know
Sending a request and reading JSON is one thing. Designing an API that other people can rely on is something where things get complicated.
Sep 19
•
ByteByteGo
277
11
Migrations at Scale: Changing the Application Engine at 30,000 Feet
In this article, we will look at how migrations work at scale and the key strategies that can help make it as efficient as possible.
Sep 17
•
ByteByteGo
68
1
How LLMs Can Find a Needle in a Haystack
In this article, we are going to look at how LLMs can find a needle in a haystack.
Sep 16
•
ByteByteGo
279
3
9
LAST CALL FOR ENROLLMENT: Build with Claude Code
We’re relaunching Build with Claude Code, a 2-day intensive cohort-based course taught by John Kim, who has trained hundreds of engineers at Meta to use…
Sep 15
•
ByteByteGo
171
3
2
Do LLMs Have the Memory of a Goldfish?
In this article, we will learn how LLMs handle memory so that they are useful to end users in performing complex tasks that require conversation and…
Sep 15
•
ByteByteGo
276
4
7
LLMs as a Judge: How to Know if Your LLM is Healthy
In this article, we are going to look at the process of LLM evaluation in detail.
Sep 14
•
ByteByteGo
278
4
7
EP225: Why Does Git Revert Cause Conflicts?
git revert looks straightforward until it throws a conflict
Sep 12
•
ByteByteGo
157
1
4
Learn Claude Code, evals, AI systems, and more: ByteByteGo Live is here
Most online courses never get finished (~4% completion). Live cohorts get ~40%, roughly 10x higher. Live courses are the only courses people actually…
Sep 11
•
ByteByteGo
189
3
6
A Guide to Application Networking Basics
In this article, we will look at the various aspects of networking in detail.
Sep 10
•
ByteByteGo
114
8
How Smart Model Routing Can Cut LLM Costs by 10X
Cost reduction isn’t a given. It also depends on the types of requests the application receives, the price difference between models, and how well the…
Sep 9
•
ByteByteGo
271
7
13
Built for Reliability: How American Express Processes Payments at Scale
In this article, we will try to understand how the transaction runs through such a cell-based architecture and how the payments are processed even when…
Sep 8
•
ByteByteGo
247
3
7
This site requires JavaScript to run correctly. Please
turn on JavaScript
or unblock scripts