成为VIP会员查看完整内容
VIP会员码认证
首页
主题
会员
服务
注册
·
登录
GPUs
关注
0
综合
百科
VIP
热门
动态
论文
精华
Accelerating Disaggregated RL for Visual Generative LLMs with Diffusion-Based Parallelism and Trainer-Assisted Generation
Arxiv
0+阅读 · 6月23日
MTGenRec: An Efficient Distributed Training System for Generative Recommendation Models in Meituan
Arxiv
0+阅读 · 6月22日
Randomized Sketching is Robust to Low-Precision Rounding on GPUs
Arxiv
0+阅读 · 6月18日
BatchGen: An Architecture for Scalable and Efficient Batch Inference
Arxiv
0+阅读 · 6月19日
Efficient Domain Decomposition for the Helmholtz Equation on GPUs
Arxiv
0+阅读 · 6月19日
Superhuman AI for Generals.io Using Self-Play Reinforcement Learning
Arxiv
0+阅读 · 6月22日
A comprehensive study on ILP acceleration accounting for sparsity, area, energy, data movement using near-memory architecture
Arxiv
0+阅读 · 6月20日
Concordia: JIT-Compiled Persistent-Kernel Checkpointing for Fault-Tolerant LLM Inference
Arxiv
0+阅读 · 6月22日
A New Sparse Matrix Vector Multiplication GPU Algorithm Designed for Finite Element Problems
Arxiv
0+阅读 · 6月17日
Performance Analysis of Digital Processing-in-Memory through a Case Study on Convolutional-Neural-Network Acceleration
Arxiv
0+阅读 · 6月18日
UltraQuant: 4-bit KV Caching for Context-Heavy Agents
Arxiv
0+阅读 · 6月19日
Private Iris Recognition with High-Performance FHE
Arxiv
0+阅读 · 6月17日
TurboServe: Serving Streaming Video Generation Efficiently and Economically
Arxiv
0+阅读 · 6月17日
UltraEP: Unleash MoE Training and Inference on Rack-Scale Nodes with Near-Optimal Load Balancing
Arxiv
0+阅读 · 6月18日
ShuntServe: Cost-Efficient LLM Serving on Heterogeneous Spot GPU Clusters
Arxiv
0+阅读 · 6月17日
参考链接
提示
微信扫码
咨询专知VIP会员与技术项目合作
(加微信请备注: "专知")
微信扫码咨询专知VIP会员
Top