Frontier Inference Clusters

San Jose
Joined April 2024
Include

Only show posts containing:

Exclude

Hide posts containing:

Time range
-
Minimum likes
If you're excited to tackle the challenges required to run the world's inference like: - Building a Gigawatt-scale factory - Orchestrating quadrillion-token clusters - Making self-improving kernel agents - Designing new product lines from scratch You should join us: etched.com/join
5
7
1
218
91,403
We're grateful for our customers, suppliers, and investors for enabling us to ramp production. Under 1% of the world has access to frontier models. Scaling intelligence requires a new kind of inference hardware. Our mission has never been more urgent. etched.com/progress/from-zer…
4
8
275
121,001
Before leading the round, Jane Street tested our hardware. In their words: “We tested the chip and are pleased with the early results. Etched’s unique approach to inference delivers the precision we will need to support our most demanding workloads.”
6
12
9
512
262,738
We've raised $700M at a $21B valuation from Jane Street, Kleiner Perkins, Sequoia, A16Z, Peter Thiel, BCV, and Blackstone. We're also excited to share that we've shipped our first rack to Jane Street.
306
452
391
6,977
7,825,249
So excited, so much still to build
10
1,824
Replying to @bhorowitz
Excited to partner with you and the a16z team!
18
1,483
Chip design is what connects our inference system. Our entire silicon flow is custom for inference. We develop critical IP in-house, including our fabrics, interconnects, math blocks, and memories. Closing blocks at low voltages is a very hard PD problem. Etched heavily invests in owning our own physical design, creating novel solutions to solve challenging problems like IR drop. We build custom PnR and sign-off flows to push PPA to its limit. Chip designers own a block through this entire process: uArch, RTL, DV, PD, signoff, and (most importantly) bringup and production!
2
3
94
55,777
Firmware at Etched builds the software that controls our inference clusters. That means custom algorithms and schemas to achieve maximum performance: doorbell-free submission paths, high-speed swapping to live spares, and nanosecond-accurate profiling tools. Firmware co-designs directly with our inference and ASIC teams to keep the lowest levels of the stack optimized for production workloads. Firmware owns the code closest to the metal and every cycle you save pushes the Pareto frontier of inference.
1
1
87
54,378
Electrical engineering at Etched is uniquely challenging. Enabling low-voltage inference and cluster-scale memory requires breakthroughs across our platform. Our PCBs handle tons of current and high-speed SerDes while being manufacturable at scale. This means asymmetric PCB stackups, vertical power delivery, high-fidelity FEA sims, and well-designed test points. EEs own our platform work from stack-up to bring-up, working closely with our firmware, thermal, and inference teams to get production workloads running as quickly as possible.
1
3
117
72,420
Inference at Etched is much more than writing kernels. We own the full stack, from KV cache management to sharding algorithms to workload-aware SerDes tuning. We push the Pareto curve daily, co-design models with customers, collaborate with chip designers, and are always improving our model harness for recursive kernel development. If you're excited by things like KV tree masking, diffusion speculators, latent compression, and large scale-up domains, Etched is the right place for you. We have an unlimited token budget.
4
6
1
185
100,437
We're hiring across inference, firmware, platform, chip design, and more. We believe in end-to-end ownership that spans beyond most engineering roles. We look for people that are deeply curious to understand the entire inference system, often beyond what's comfortable.
4
4
256
148,572
We're grateful for the support from our partners, suppliers, and team. We have a lot of work to do as we ramp production. etched.com/progress
6
3
1
329
260,192
We’ve raised $300M in Series C funding at a $10.3B valuation from Sequoia, Andreessen Horowitz, Jane Street, Argo, and SK Hynix. Our mission is to run the world's inference. This round accelerates production of our inference clusters. We've opened an 80,000-sqft, 10-MW facility 15 minutes from our office to expedite production and prototyping.
183
355
223
5,573
7,535,400
Vertically integrating the full inference hardware stack is extremely hard. Hear @UbertiGavin and @robertwachen share three years of battle stories leading into today's launch
We're coming out of stealth. We've built our first racks after a successful A0 tapeout, $1B+ in customer contracts, and $800m raised. Early customer tests show us achieving SOTA throughput, latency, and power efficiency on inference workloads. Our first racks ship this summer.
25
61
17
608
291,366
Replying to @LiamFedus
Thanks for all of your support Liam!
18
17,257
Replying to @karpathy
We're excited to spend more time together!
2
88
14,480
Replying to @karimatiyeh
More to come :)
3
1
24
28,720
Thanks Tim, you're always welcome!
1
15
32,194