2026-d1-1200-sponsor-talk-meta

Archived conversation · Aug 18, 2026 1:17 PM – Aug 20, 2026 12:58 AM · 40 messages
Tuesday, August 18, 2026

Ryan Scherbarth (nvidia) joined the channel
David Ozog joined the channel
Sayan Ghosh joined the channel
Jiaqi Lou joined the channel
bsahu joined the channel
Fabian joined the channel
Francois Labonte joined the channel
Daud Arslan joined the channel
Shahbozbek Hakimov joined the channel
Tong Xu joined the channel
myoung-gyun.suh joined the channel
xi.chen joined the channel
Mahfuz joined the channel
Varun Chotalia joined the channel
pierre-louis.benard joined the channel
Matthew Fricke joined the channel
Advm joined the channel
ahmadrazarehman033 joined the channel
Howard Wang joined the channel
rashid-ahmed.kukkady joined the channel
Wednesday, August 19, 2026

L
Leila Rashidi 2:06 PM
How CPUs are used with MTIA for agentic AI usecases?
1 reply
K
kirtesh 2:47 PM
Unfortunately I can't share details here because this is part of currently active discussions.
M
Matthew Fricke 2:09 PM
Are those RDMA NICs infiniband or ethernet? If ethernet RoCE or something else. .
3 replies
L
Leila Rashidi 2:13 PM
It was roce
✅ 1
K
kirtesh 2:31 PM
Yes, these are RDMA / RoCE NICs.
L
Leila Rashidi 2:32 PM
Has Meta any plan to use UFH rather than ESUN 1.0 header?
H
Hesham ElBakoury 2:13 PM
what is difference between MTIA and GPU
1 reply
K
kirtesh 2:43 PM
MTIA is Meta's custom line of accelerators that are designed specifically for for Ranking & Recommendation (R&R) and GenAI workloads. Resulting in similar or higher performance and cost efficiency. See ai.meta.com/blog/meta-mtia-scale-ai-chips-for-billions
A
ashkan.sobhani 2:13 PM
Could you provide some examples where this fungibility for scale-up vs scale-out helps?
3 replies
K
kirtesh 2:48 PM
As mentioned in the slides, we could potentially use all the NICs for scale-up connectivity depending on the workload requirements.
👍 1
A
ashkan.sobhani 3:03 PM
Thanks for the response, do you have any insights on workload requirements in terms of scale-up/scale-out ratio?
K
kirtesh 12:58 AM
Generally speaking training requires collectives over accelerators in multiple racks and inference jobs tend to be smaller in scale resulting in different requirements.
M
Mike Capuano 2:14 PM
What SerDes are you using?
1 reply
K
kirtesh 2:54 PM
I'm unable to share this information at this time.
R
Ramtin Soleymani 2:15 PM
Is the choice between Message Engine offload and direct PE injection made dynamically based on message size or latency, or is it decided ahead of time in software?
3 replies
L
Leila Rashidi 2:19 PM
You may find your response in MITA paper accepted for super computing 2026. Arxiv version is available
R
Ramtin Soleymani 4:18 PM
Thank you Leila for the response
K
kirtesh 12:50 AM
I'd also recommend arxiv.org/html/…. As for the question - typically its a decision made up front on the host side. However, dynamic decision in device is also possible and something we are exploring.
L
Leila Rashidi 2:43 PM
Is it possible to converge scale up and out in future? Does Meta experience reduction of scale up bandwidth to scale out bandwidth ratio? 5 is much lower than 10, which has been observed in industry in the past
L
Leila Rashidi 2:56 PM
Is ESUN header used for scale up? Is there any header optimization for scale up? Are you running RoCE over ESUN?