Check out the latest work on "Cost-Effective Extension of DRAM-PIM for Group-Wise LLM Quantization" by Byeori Kim, Changhun Lee, Gwangsun Kim, Eunhyeok Park.
https://t.co/oVINrjdIsw
Check out the latest work on "Optically Connected Multi-Stack HBM Modules for Large Language Model Training and Inference" by Yanghui Ou, Hengrui Zhang, Austin Rovinski, David Wentzlaff, Christopher Batten.
https://t.co/kG2YH36QP7
Check out the latest work on "Comprehensive Design Space Exploration for Graph Neural Network Aggregation on GPUs" by Hyunwoo Nam, Jay Hwan Lee, Shinhyung Yang, Yeonsoo Kim, Jiun Jeong, Jeonggeun Kim, Bernd Burgstaller.
https://t.co/61Akej46PA
Check out the latest work on "RoPIM: A Processing-in-Memory Architecture for Accelerating Rotary Positional Embedding in Transformer Models" by Yunhyeong Jeon, Minwoo Jang, Hwanjun Lee, Yeji Jung, Jin Jung, Jonggeon Lee, Jinin So, Daehoon Kim.
https://t.co/8l514VbFeU
Check out the latest work on "Accelerating Page Migrations in Operating Systems With Intel DSA" by Jongho Baik, Jonghyeon Kim, Chang Hyun Park, Jeongseob Ahn.
https://t.co/6tScHii3kH
Check out the latest work on "GPU-Centric Memory Tiering for LLM Serving With NVIDIA Grace Hopper Superchip" by Woohyung Choi, Jinwoo Jeong, Hanhwi Jang, Jeongseob Ahn.
https://t.co/rf2dOmYKsI
Check out the latest work on "Cooperative Memory Deduplication With Intel Data Streaming Accelerator" by Houxiang Ji, Minho Kim, Seonmu Oh, Daehoon Kim, Nam Sung Kim.
https://t.co/RkZ12ydjs0
Check out the latest work on "SPAM: Streamlined Prefetcher-Aware Multi-Threaded Cache Covert-Channel Attack" by E. Kritheesh, Biswabandan Panda.
https://t.co/U5XXQS2pI1
Check out the latest work on "High-Performance Winograd Based Accelerator Architecture for Convolutional Neural Network" by Vardhana M, Rohan Pinto.
https://t.co/o16Tba12LM
Check out the latest work on "PINSim: A Processing In- and Near-Sensor Simulator to Model Intelligent Vision Sensors" by Sepehr Tabrizchi, Mehrdad Morsali, David Pan, Shaahin Angizi, Arman Roohi.
https://t.co/5ruBdh5zrf
Check out the latest work on "Characterization and Analysis of the 3D Gaussian Splatting Rendering Pipeline" by Jiwon Lee, Yunjae Lee, Youngeun Kwon, Minsoo Rhu.
https://t.co/E4g3vsUlAM
Check out the latest work on "Electra: Eliminating the Ineffectual Computations on Bitmap Compressed Matrices" by Chaithanya Krishna Vadlamudi, Bahar Asgari.
https://t.co/TpvLJrflTd
Check out the latest work on "Straw: A Stress-Aware WL-Based Read Reclaim Technique for High-Density NAND Flash-Based SSDs" by Myoungjun Chun, Jaeyong Lee, Inhyuk Choi, Jisung Park, Myungsuk Kim, Jihong Kim.
https://t.co/GvMnHiVpSU
Check out the latest work on "IntervalSim++: Enhanced Interval Simulation for Unbalanced Processor Designs" by Haseung Bong, Nahyeon Kang, Youngsok Kim, Joonsung Kim, Hanhwi Jang.
https://t.co/FEPpPcttaS
Check out the latest work on "Quantum Assertion Scheme for Assuring Qudit Robustness" by Navnil Choudhury, Chao Lu, Kanad Basu.
https://t.co/OmKSdYOMis
Check out the latest work on "SCALES: SCALable and Area-Efficient Systolic Accelerator for Ternary Polynomial Multiplication" by Samuel Coulon, Tianyou Bao, Jiafeng Xie.
https://t.co/aR8myr1ThK
Check out the latest work on "GCStack: A GPU Cycle Accounting Mechanism for Providing Accurate Insight Into GPU Performance" by Hanna Cha, Sungchul Lee, Yeonan Ha, Hanhwi Jang, Joonsung Kim, Youngsok Kim.
https://t.co/vzImNFBces
Check out the latest work on "A Case for Hardware Memoization in Server CPUs" by Farid Samandi, Natheesan Ratnasegar, Michael Ferdman.
https://t.co/cxkXSszP8g
Check out the latest work on "Characterization and Analysis of Text-to-Image Diffusion Models" by Eunyeong Cho, Jehyeon Bang, Minsoo Rhu.
https://t.co/sd15sBaP4d