All Stories

  1. WISP: Waste- and Interference-Suppressed Distributed Speculative LLM Serving at the Edge via Dynamic Drafting and SLO-Aware Batching
  2. WISP: Waste- and Interference-Suppressed Distributed Speculative LLM Serving at the Edge via Dynamic Drafting and SLO-Aware Batching
  3. WISP: Waste- and Interference-Suppressed Distributed Speculative LLM Serving at the Edge via Dynamic Drafting and SLO-Aware Batching
  4. SLED: A Speculative LLM Decoding Framework for Efficient Edge Serving
  5. DeepPUFSCA: Deep learning for Physical Unclonable Function attack based on Side Channel Analysis support
  6. Relax and don't Stop: Graph-aware Asynchronous SSSP
  7. Exploiting Data Redundancy in CKKS Encoding for High-Speed Homomorphic Encryption
  8. Differentiating Set Intersections in Maximal Clique Enumeration by Function and Subproblem Size
  9. Resource-Efficient Convolutional Networks: A Survey on Model-, Arithmetic-, and Implementation-Level Techniques
  10. How to design custom floating-point formats and emulate them efficiently on standard hardware
  11. MASTIFF
  12. LOTUS
  13. Exploiting in-Hub Temporal Locality in SpMV-based Graph Processing
  14. Making Consensus Less Writey to Increase Throughput
  15. Graph analytics can be accelerated using processors' vector instructions such as AVX-512
  16. Online estimation of scalability of multi-threaded programs, with minimal runtime overhead
  17. A programming model and runtime system for significance-aware energy-efficient computing
  18. A significance-driven programming framework for energy-constrained approximate computing
  19. Software-managed energy-efficient hybrid DRAM/NVM main memory
  20. A programming model and runtime system for significance-aware energy-efficient computing
  21. Analysis of dependence tracking algorithms for task dataflow execution