Fireworks AI
Fireworks is the fastest way to build, tune, and scale AI on open models. Ship production-ready AI in seconds on our globally distributed cloud infrastructure, optimized for your use case. Fireworks powers production workloads at companies like Uber, Doordash, Notion, and Cursor—delivering 15× faster speed, 4× lower latency, and 4× more concurrency than closed models.www.fireworks.ai
37
MENTIONS
8
EPISODES
6
PODCASTS
Search complete. 37 mentions across 8 episodes found for "Fireworks AI".
Sep 21, 2026
20VC: Why AI Cannot Replace Humans in Enterprise | Why Work Processes Not Models Will Be The Most Valuable Asset in AI | Why Europe Has Lost and Building in the US vs EU with Daniel Dines, UiPath
H
47:34Harry StebbingsHOST
And yeah, I think that goes to the statement of kind of owning your own intelligence, not renting it.
H
47:37Harry StebbingsHOST
That's why we invested in Fireworks, and I believe in the open model ecosystem.
D
47:41Daniel DinesGUEST
Yeah.
D
47:41Daniel DinesGUEST
You know, I'm a big fan of Fireworks, and we are using them, uh, quite a bit.
H
47:46Harry StebbingsHOST
Do you like them?
D
47:47Daniel DinesGUEST
Yes.
D
48:32Daniel DinesGUEST
The real investment for an enterprise is to creating this map of work that documents how they actually work.
D
48:39Daniel DinesGUEST
And with this one, they can train their own models.
The current balance of power in open models
N
9:13Nathan LambertHOST
These open platforms are the best approximation of open model usage we have.
N
9:17Nathan LambertHOST
A large proportion of open model usage is on platforms that do not disclose the per model breakdowns, such as Together AI or Fireworks AI, and in private deployments for enterprise applications.
N
9:28Nathan LambertHOST
Many prominent technology companies and startups have been building on Chinese open weight models for their AI features, such as Harvey, the legal startup, Cursor, the coding agent, and DoorDash's usage of Kimi models.
N
9:40Nathan LambertHOST
Airbnb's use of Quen, or Perplexity's use of DeepSeek.
RevOps for Hypergrowth (Fireworks AI, Perplexity, Exa & More) | CEO @ Go Nimbly, Jen Igartua
S
0:00speaker_0NARRATOR
Jen Igartwa is the founder and CEO of GoNimbly, a revenue operations firm that has built the go-to-market engines for SaaS unicorns like Zendesk, Twilio, and Gong, as well as some of the fastest-growing AI companies.
J
0:13Jen IgartuaGUEST
I'm working now with Whisperflow, Exa, Perplexity, and Fireworks, these AI-first companies.
J
0:19Jen IgartuaGUEST
The number of people they want to hire in a year is like, it blows my mind.
S
0:22speaker_0NARRATOR
In today's episode, Jen shares how to prioritize when your customer demand outpaces your sales infrastructure, and her favorite approach if you want to get ROI from a RevOps team fast.
S
1:43Sam JacobsHOST
She is great on screen, on camera.
S
1:46Sam JacobsHOST
She hosts This Week in SaaS, which you can catch on LinkedIn.
S
1:50Sam JacobsHOST
And today, GoNimbly is servicing some of the fastest growing AI native companies in the world, including companies like Fireworks.
S
1:57Sam JacobsHOST
Jen is also the co-founder of Pillbox Games, an independent board game company.
20VC: How to Build Your Own Data Center & Why Every Startup Should Do It | How ElevenLabs Leapfrogged Us: What I Learned | The AI Talent War: How Your Hiring Process Needs to Change with Cliff Weitzman, Speechify
C
21:02Cliff WeitzmanGUEST
We have, but like small data sets.
H
21:04Harry StebbingsHOST
So I suggest you use Fireworks.
C
21:06Cliff WeitzmanGUEST
Okay.
H
21:06Harry StebbingsHOST
But, uh, I mean, Fireworks is amazing.
H
21:08Harry StebbingsHOST
Linda, the founder is, she one of the co-founders of PyTorch.
H
21:11Harry StebbingsHOST
But I, I, I had the very obvious realization that you'd have every company having their own specialized models of a certain size trained on their own data, but you would need supplemental data-
8 MINS LATER
H
29:34Harry StebbingsHOST
Value accrues to top one player.
C
29:36Cliff WeitzmanGUEST
I, I agree.
What Precision Is Your Model Actually Running At?
H
0:43Herman PoppleberryHOST
The model card says the weights are stored at a certain precision, FP16, BF16, whatever the published architecture specifies.
H
0:52Herman PoppleberryHOST
But what precision are they actually running at when your request lands? Daniel wants to know whether closed source vendors quantize their own flagship models, whether commercial inference providers like Together and Fireworks run everything at full precision, whether unofficial quants get deployed when demand spikes, and whether your customer tier determines whether you get the quantized version or the full one.
H
1:15Herman PoppleberryHOST
Basically, is the model you think you're calling the model you're actually
C
1:19CornHOST
getting? And the short answer is...
C
1:53CornHOST
On one side, you've got the closed-source vendors, OpenAI, Anthropic, Google.
C
1:59CornHOST
They train the models, they host them, they control the whole stack.
C
2:03CornHOST
Then you've got the commercial inference platforms, Together, AI, Fireworks AI, Grok, Replicate.
C
2:10CornHOST
These are companies whose whole business is hosting models, mostly open-weight ones, and serving them through an API.
Runway's Solaris Renders Websites as Live Video With No Code Underneath
S
13:53speaker_0HOST
Let's run through those quick hits.
S
13:55speaker_0HOST
First, we have Fireworks, which is a platform focused on training open models cheaper and faster for these specific workflows.
S
14:02speaker_1HOST
Essential infrastructure.
S
14:03speaker_0HOST
Then there is OpenClaw 2.0, which is an open source personal agent that just added multiplayer sessions.
20VC: Inside Sequoia's Investment Committee: Lessons from Don Valentine, Doug Leone and Alfred Lin | How the SpaceX and Citadel Deals Went Down | What Sequoia Specifically Looks for in Founders with Julien Bek
J
53:53Julien BekGUEST
Mm-hmm.
H
53:54Harry StebbingsHOST
And I feel a lot more certainty when I invest in Fireworks, when I invest in Macaw, when I invest in ClickHouse-
J
54:01Julien BekGUEST
Mm-hmm
H
54:01Harry StebbingsHOST
...
H
54:01Harry StebbingsHOST
the infrastructure that I know whoever wins at the top layer in the app layer wins.
H
54:06Harry StebbingsHOST
But they're gonna use Fireworks, they're gonna use ClickHouse, they're gonna use Macaw to get there.
H
54:11Harry StebbingsHOST
Do you not just sit around the table as a partnership and go, "God, the infrastructure layer is much easier and better."
J
54:16Julien BekGUEST
[laughs]
The Winchester Mystery House Problem in AI Development
D
25:57Drew BreunigGUEST
...
D
25:57Drew BreunigGUEST
I'm using Claude and I'm using Fireworks serverless.
D
26:01Drew BreunigGUEST
Um, the other reason I like Fireworks is they have good data, um, agreements.
D
26:06Drew BreunigGUEST
They aren't, they aren't training off my data, they aren't retaining my data, um, and so I can make those decisions.
D
26:11Drew BreunigGUEST
So like when K3 came out, I've been a Kimi fanboy since K2.
D
26:15Drew BreunigGUEST
It's one of my favorite system papers ever is K2.
D
26:19Drew BreunigGUEST
Um, and, uh, but I had to wait for it to, to roll out to Fireworks, um-
D
26:26Demetrios BrinkmannHOST
Yeah