Jiaxun Han

According to our database1, Jiaxun Han authored at least 3 papers between 2024 and 2026.

Collaborative distances:
  • Dijkstra number2 of four.
  • Erdős number3 of four.

Timeline

Legend:

Book  In proceedings  Article  PhD thesis  Dataset  Other 

Links

On csauthors.net:

Bibliography

2026
Latency-SLO-Aware Memory Offloading for Large Language Model Inference.
Proceedings of the 40th ACM International Conference on Supercomputing, 2026

2025
Memory Offloading for Large Language Model Inference with Latency SLO Guarantees.
CoRR, February, 2025

2024
OneGraph: a cross-architecture framework for large-scale graph computing on GPUs based on oneAPI.
CCF Trans. High Perform. Comput., April, 2024


  Loading...