Stockholm, Sweden

#39 - Model Optimization & CPU Based Inferencing

Europe/Stockholm (GMT+2)
until 21:00
AI Sweden · Stockholm

#39 - Model Optimization & CPU Based Inferencing takes place on Thu, 24 Sept 2026 at 17:00 (GMT+2) at AI Sweden in Stockholm, Sweden, and runs until 21:00. Entry is free; the listing is on Luma.

About this event

All right!!! Meetup #39 will take place September 24 at AI Sweden and focus on "Model Optimization & CPU Based Inferencing". We're super excited to team up with Spanish Quantum AI wizards Multiverse Computing and tech giants Intel and HPE. The program is currently being worked out, so stay tuned for updates!

Read full description

There is a tremendous amount happening in the area of model optimization and inferencing. Novel approaches to both seem to be announced daily, such as making it possible to run the recently released 2.78 trillion parameter model Kimi K3 on a single CPU with 8GB of memory...!!! or Multiverse Computing's July 23 announcement that all their compressed models now run on Intel Xeon 6 Processors. So buckle up and brace for impact! Event Program- Doors open at 17:00 CET- Talks begin at 17:45 CET- There will be pizzas & drinks- There may be a moderated Q&A session.... Speaker Line-Up TBD... - Franco Serra, Solutions Architect at Multiverse Computing will give a talk titled "Ultra Efficient Models to Scale your GenAI Datacenter & Fit for purpose on the Edge deployment". Abstract: As organizations race to deploy Generative AI, two challenges dominate: how to scale inference economically in the datacenter, and how to bring intelligence on the edge where connectivity, power and footprint are constrained. Multiverse Computing makes ultra-efficient, compressed AI models, including LLMs, VLMs, speech-to-text, and computer vision models — engineered to deliver the same accuracy with a fraction of the memory requirements. By integrating with Intel® hardware, enterprises can deploy larger AI models on existing infrastructure rather than expanding it. In the datacenter, leaner models translate directly into higher throughput, lower energy consumption, and improved ROI by enabling more users and workloads to run on the same infrastructure. At the edge, the same compression breakthroughs make it possible to deploy GenAI in constrained environments — bringing secure, low-latency AI to tactical environments. - Jonas Svennebring, Principal Engineer at Intel and Theo Charitidis, Machine Learning Engineer also at Intel will give a talk titled "Architecting the future AI/ML CPUs". Abstract: The future of AI/ML capable CPUs will depend on architectures that balance computational performance, memory efficiency, and adaptability. To that end, it is essential that high multiply-accumulate (MAC) throughput enabled, supported by a well-balanced memory subsystem. Another key architectural guidance is figuring out which workloads should realisticly run on general-purpose cores and which should be delegated to specialized accelerators. Finally, supporting the correct dynamic range and numerical precision will be critical for striking the right balance between reliable application results and minimizing hardware costs." - Johan Fondin, Solution Architect at HPE will give a talk will give a talk titled "Out of the box KV-cache acceleration". Abstract: How do you increase the number of concurrent sessions on a GPU by a hundredfold, and 20x times faster as well? Learn how HPE solves this issue with their new KV-cache acceleration technology for AI platforms. Looking forward to seeing you there, /Patrick & the Stockholm MLOps Team

Tickets
Luma
Organizer
Stockholm MLOps
Time zone
GMT+2
Updated

Similar events

View all

Details

When
Thu, 24 Sept 2026 · 17:00 (GMT+2) · until 21:00
Where
AI Sweden
Address
AI Sweden, Folkungagatan 44, 118 26 Stockholm, Sweden
Price
Free
Genre
ai

Questions

When is #39 - Model Optimization & CPU Based Inferencing?
Thu, 24 Sept 2026 at 17:00 GMT+2.
How much are tickets for #39 - Model Optimization & CPU Based Inferencing?
Entry is free.
Where is #39 - Model Optimization & CPU Based Inferencing?
AI Sweden, AI Sweden, Folkungagatan 44, 118 26 Stockholm, Sweden, Stockholm, Sweden.
Where can I buy tickets for #39 - Model Optimization & CPU Based Inferencing?
Tickets are sold on Luma. This page links straight to that listing; no tickets are sold here.
What time does #39 - Model Optimization & CPU Based Inferencing end?
It runs until 21:00.

Keep browsing