Jun 22, 2026 · 40 min · 12 segments
This week we welcome American Association of Law Libraries leaders Jenny Foster, AALL President for 2025-2026, and…
Jenny FosterGuest
Jessica WhytockGuest
Greg LambertHost
Marlene GebauerHostSam MooreGuest
Hi, I'm Marlene Gebauer from The Geek in Review, and I have Sam Moore here from Legal Technology Hub, who's gonna tell us a little bit about analysis of token usage and model selection.
Well, uh, it is tokens, tokens everywhere, I think spurred on by the launch of Claude for Legal, but certainly going back further than that.
And the topics we're discussing with clients right now tend to fall into three interconnected topics.
First is model selection, because a lot of these products give the users a choice of which model they want to use for a given prompt.
I've seen law firm clients whose users just pick the most sophisticated model for everything, toggle on every optional feature available, and then are confused as to why responses are taking a long time and why they're hitting token limits very, very quickly.
I've had several conversations lately about what a context window even is, um, how it can create drift when it gets crowded in a chat's context window, and why that really matters for legal use cases, which often involve uploading quite large documents, which take up a lot of space in those context windows.
Law firms and law departments, I think, are generally not that accustomed to this kind of pay-as-you-go model in technology, not unless you're like me and you recall when the big legal research platforms were on a pay-per-search basis.
So now those users are running into high-cost overages on the frontier models in particular, and they're realizing that low sticker price per month is not their reality, not when their users don't know how to use those tools efficiently and how to control cost.
So as well as delivering advisory work on these topics on a one-to-one basis, we're actually working on a series of articles for LTH Premium about these topics, which we'll then combine into a sort of playbook for our subscribers to keep handy when they're working with GenAI tools.
And if people want to know more about LTH Advisory and what we can do, they can always get in touch with us by going to legaltechnologyhub.com or by finding me on LinkedIn.
Read the full transcript.
Create an account to read the whole episode, search across every transcript, and follow the shows you care about.

Hi, I'm Marlene Gebauer from The Geek in Review, and I have Sam Moore here from Legal Technology Hub, who's gonna tell us a little bit about analysis of token usage and model selection.
Well, uh, it is tokens, tokens everywhere, I think spurred on by the launch of Claude for Legal, but certainly going back further than that.
And the topics we're discussing with clients right now tend to fall into three interconnected topics.
First is model selection, because a lot of these products give the users a choice of which model they want to use for a given prompt.
I've seen law firm clients whose users just pick the most sophisticated model for everything, toggle on every optional feature available, and then are confused as to why responses are taking a long time and why they're hitting token limits very, very quickly.
I've had several conversations lately about what a context window even is, um, how it can create drift when it gets crowded in a chat's context window, and why that really matters for legal use cases, which often involve uploading quite large documents, which take up a lot of space in those context windows.
Law firms and law departments, I think, are generally not that accustomed to this kind of pay-as-you-go model in technology, not unless you're like me and you recall when the big legal research platforms were on a pay-per-search basis.
So now those users are running into high-cost overages on the frontier models in particular, and they're realizing that low sticker price per month is not their reality, not when their users don't know how to use those tools efficiently and how to control cost.
So as well as delivering advisory work on these topics on a one-to-one basis, we're actually working on a series of articles for LTH Premium about these topics, which we'll then combine into a sort of playbook for our subscribers to keep handy when they're working with GenAI tools.
And if people want to know more about LTH Advisory and what we can do, they can always get in touch with us by going to legaltechnologyhub.com or by finding me on LinkedIn.