| |
|
| |
Inference is the critical step that shapes the speed, cost, & quality of every request your application makes. With new models releasing weekly, the infrastructure landscape is moving just as fast as the models themselves.
Choosing an inference provider has many considerations including, speed, cost (incl. cache pricing discounts), latency, & day zero support for new models.
Artificial Analysis benchmarks over 500 inference endpoints across speed, cost, latency, & quality. Our evaluations show output speeds for the same model vary by 15x across providers. For the same query, end-to-end response time can vary by up to 12, while the cost difference between the cheapest & most expensive provider can exceed 3x.
Join the Artificial Analysis team for an evening with leaders across AI inference. We'll examine the key inference trends, unpack what the benchmark data reveals, & discuss what these differences mean for those building with AI.
Agenda:
6:00 PM: Doors open
6:30 PM: Artificial Analysis talk
7:00 PM: Panel
7:30 PM: Networking, light bites, & drinks
8:30 PM: Event ends
Speakers:
To be announced soon
Photography & videography disclaimer: By attending this event, you acknowledge that photography & videography may take place. If you do not consent, please contact the event host.
|
|
|
|
|
|
|
|