SIGN IN SIGN UP

llama-mmap : avoid a second full-size copy of each tensor with direct-io (#29749)

Assisted-by: Claude

Co-authored-by: Pranesh Gonegandla <pgonegandla@nvidia.com>
P
Pranesh Gonegandla committed
32dd62ee6dfa80ada846551fefec215cefc5ae1c
Parent: f11d642
Committed by GitHub <noreply@github.com> on 10/1/2026, 7:44:31 AM