2026-d2-1000-sponsor-marvell

Archived conversation · Aug 18, 2026 1:20 PM – Aug 20, 2026 3:16 PM · 56 messages
Tuesday, August 18, 2026

Ryan Scherbarth (nvidia) joined the channel
David Ozog joined the channel
Jiaqi Lou joined the channel
bsahu joined the channel
Fabian joined the channel
Francois Labonte joined the channel
Daud Arslan joined the channel
Shahbozbek Hakimov joined the channel
Tong Xu joined the channel
myoung-gyun.suh joined the channel
xi.chen joined the channel
Mahfuz joined the channel
Varun Chotalia joined the channel
pierre-louis.benard joined the channel
Matthew Fricke joined the channel
Advm joined the channel
ahmadrazarehman033 joined the channel
Howard Wang joined the channel
rashid-ahmed.kukkady joined the channel
Kulika Weizman joined the channel
Thursday, August 20, 2026

R
rmahatme363 11:30 AM
Hello & Good Morning, This is Ravi Mahatme from Marvell.
J
jay.gill 12:01 PM
Please ask Ravi to share slides
C
Chris Browning (Black Semi) 12:01 PM
Is he sharing?
A
ashkan.sobhani 12:01 PM
nothing is shared
P
pierre-louis.benard 12:02 PM
no
S
Sayan Ghosh 12:02 PM
we will inform Ravi
šŸ™Œ 1
L
Leila Rashidi 12:11 PM
What about software support for using memory appliance?
1 reply
R
rmahatme363 12:18 PM
This works as a standard CXL driver. For integration with AI libraries, this will need integration with AI libraries like vLLM, LMCache. This can be done by customers or we have a team looking into it as well
L
Leila Rashidi 12:12 PM
Does PFMA support 2 tier scale up?
1 reply
R
rmahatme363 12:19 PM
Yes it has the capability to be tiered to increase the radix but it adds a hop
M
Mike Capuano 12:12 PM
What company is building your shuffle or is this designed and built in house?
1 reply
R
rmahatme363 12:19 PM
This is designed in house by our photoncis team but we work with external vendors to have this manufactured
J
jay.gill 12:14 PM
Do you have plans to increase the fan-out connectivity from this appliance to more than 16, e.g. by introducing a switching fabric?
1 reply
R
rmahatme363 12:22 PM
There are multiple ways we can increase the radix
1. Add more appliances as a 2nd tier but ths will add a hop
2. Add an OCS as the 2nd tier, this probably is the lowest latency solution
3. Build a 2nd generation appliance which has a higher radix.
What we have right now is our Gen1 appliance
J
Jan Gray 12:15 PM
Very interesting design. Thank you for the presentation. What is the utility of two HBM3e memories (>2 TB/s) given the remote memory access bandwidth into/out of the device seems to be 16*224Gb/s=~ 448 GB/s? Would not one HBM3e suffice as a fast "L4 cache" in front of the fully populated dimms? (What is the approx. total DRAM bandwidth to the DDR DIMMs?)
1 reply
R
rmahatme363 12:24 PM
With 2 HBM you can have lot more transactions in flight as each HBM has 32 virtual channels. The max value of this appliance is when you integrate with the PF-chiplet as it gives you full bandwidth

For DRAM bandwidth, these run at 5200Mbps and 2 DPC
L
Leila Rashidi 12:15 PM
Why not to use PFMA as scale up switch?
1 reply
R
rmahatme363 12:25 PM
It is not a full featured scale-up switch with all advanced features a typical switch has but it can provide basic scale-up switch functionality if required.
A
Amit Jha 12:15 PM
Can we please get the slides?
2 replies
R
rmahatme363 12:25 PM
Yes, please email me : <mailto:rmahatme@marvell.com|rmahatme@marvell.com>
A
Amit Jha 12:29 PM
Thanks a lot
M
Mohammed Mahfuz 12:16 PM
What is the energy efficiency for Marvell photonic fabric in terms of pJ/bit?
4 replies
L
Leila Rashidi 12:17 PM
It is available on Celestial ai presentation
šŸ‘ 1
M
Mohammed Mahfuz 12:21 PM
Where can I find out about this presentation?
Can you share it with me? Thanks
L
Leila Rashidi 12:21 PM
Watch hot chip presentation
L
Leila Rashidi 12:21 PM
Or message me on LinkedIn. I can share
šŸ‘ 1
L
Leila Rashidi 12:19 PM
What are supply chain challenges for building PFMA?
3 replies
R
rmahatme363 12:34 PM
This is a multi-disciplinary device incorporating advanced 5nm ASICs, optical interposers, HBM , shuffles, lasers etc, so needs a well oiled supply chain.
We have been working on it since Celestial AI times and now after joining Marvell, we can leverage Marvell's supply chain expertise and resources
L
Leila Rashidi 12:38 PM
When product will be ready?
R
rmahatme363 3:16 PM
Please email me at <mailto:rmahatme@marvell.com|rmahatme@marvell.com>
S
Sayan Ghosh 12:19 PM
@Ravi Mahatme
S
Sayan Ghosh 12:20 PM
@rmahatme363 sorry for mistagging, see questions above
N
nicky 12:28 PM
That was a fascinating talk thank you Ravi, I am not sure if I am wording this well but at what point can you start see kv cache as behaving like a local memory, is there some kind of latency budget you see for it?
1 reply
R
rmahatme363 12:37 PM
I think we are already seeing KVCache change from being stateless (for basic chat applications) to being stateful (for agentic AI) , so I think there is going to be a requirement for both - lots of fast local memory ( for hot &amp; warm KVCache) and for long term storage ( to store stateful KVCache)
šŸ‘ 1šŸ™Œ 1
A
Ahmed Khalil (unaffiliated) 12:31 PM
Since Marvell also has the newer Structera S CXL switches, have you compared the appliance against a current CXL 3.x pool? The paper comparison seems to use an older CXL 2.0 switched setup, so I’m curious how much of the latency gap remains with the latest electrical baseline.
1 reply
R
rmahatme363 12:37 PM
Our perf team is working on this.