Library

Subject
Tags

3,552 matches · cs.AR

#
001ARES: Adaptive Reasoning-Effort Steering for PPA- and Cost-Aware RTL Optimization with LLM AgentsarXivPaperStef Cuyckens et al.Yesterday
002Demystifying DRAM Read Disturbance: Bridging the Gap Between Experimental Characterization and Device-Level Modeling of RowHammer and RowPress PhenomenaarXivPaperHaocong Luo et al.Yesterday
003GyRot: Leveraging Hidden Synergy between Rotation and Fine-grained Group Quantization for Low-bit LLM InferencearXivPaperSangjin Kim et al.Yesterday
004LightRot: A Light-Weighted Rotation Scheme and Architecture for Accurate Low-Bit Large Language Model InferencearXivPaperSangjin Kim et al.Yesterday
005Nanoparticle Networks for Neuromorphic ComputingarXivPaperJonas Mensing et al.Yesterday
006Investigating reservoir computing for branch predictionin pipelined processors using emerging CMOS memristor devicesarXivPaperHarvey Samuel George Johnson et al.2 days ago
007LLMET: Enabling Cross-Layer Evaluation of Emerging M3D Memories for Energy-Efficient LLM ServingarXivPaperMing-Yen Lee et al.2 days ago
008At-the-Roofline Sparse Tensor Contractions on Vector Processors for Transformer InferencearXivPaperBowen Wang et al.3 days ago
009ContractHIL-HLS: Contract-Aligned Multi-Agent Workflow with Hardware-in-the-Loop Feedback for HLS DesignarXivPaperJingbo Zhang, Haoxiang Sun et al.3 days ago
010MDTransformer: A Hardware-Software Co-Design of Mode-Division Photonic Transformer Accelerator with Inverse-Designed Coherent CrossbararXivPaperS. Serunjogi, Rachmad Vidya Wicaksana Putra et al.3 days ago
011Behavior-Driven ExplainabilityarXivPaperCaroline Dominik et al.4 days ago
012The SpiNNaker2 chip: a many-core platform for flexible and scalable brain-inspired computingarXivPaperStefan Scholze et al.4 days ago
013ADVERSARIAL: And-Inverter Graph-Assisted Hardware Trojan Detection At ScalearXivPaperYaroslav Popryho et al.5 days ago
014SPARC: Automated Root-Cause Analysis of Pre-Silicon Power Side-Channel Leakage in the Processor Design FlowarXivPaperAndrija Nešković et al.6 days ago
015FusionML: Prefill, Not Decode - Mechanism and Boundaries of CPU+GPU Co-Execution on Unified-Memory Apple SiliconarXivPaperOm Mohite7 days ago
016HiKV: Hierarchical Importance-Aware KV Cache with Hardware Acceleration for LLM DecodingarXivPaperChao Fang, Jun Yin et al.7 days ago
017Multi-primitive in-memory computing for Monte Carlo tree searcharXivPaperTergel Molom-Ochir et al.7 days ago
018Optimizing Transformer Neural Network for Real-Time Outlier Detection on FPGAsarXivPaperIlia Sobakinskikh et al.7 days ago
019Sparse by Command: Task-Conditional Compute Skipping for Multi-Task Inference AcceleratorsarXivPaperAfzal Ahmad, Gaoyu Mao et al.7 days ago
020Unified Static-Dynamic Pruning for Efficient LLM InferencearXivPaperDeyu Yang, Rundong Wei et al.7 days ago
021Benchmarking LLMs for Verilog Design FlowsarXivPaperAngshuman Chakravertty et al.Jul 23
022DRC-Aid: Design-Rule Correction via Agentic Framework utilizing Inference-Time Large Language ModelsarXivPaperAnushka Mukherjee et al.Jul 23
023Hardware-Software Co-Design for Float16 On-Device Training on RISC-V Single-CorearXivPaperBenjamin Hubinet, P. Moellic et al.Jul 23
024RED-PIM: Reducing Data Movement for Transformers using Processing-in-MemoryarXivPaperAlexandru Meterez, Pranav Ajit Nair et al.Jul 23
025AlphaRoute: Large Language Models as Semantic Optimizers for Multi-Objective RoutingarXivPaperKabir Murjani, Mishri Bhavsar et al.Jul 22
026Formal Foundations for Known Good Reliable Die Screening in Chiplet-Based AI Systems-on-ChiparXivPaperPrashanthi Metku, Chandra GanduJul 22
027Leveraging ECRAM for Edge Continual LearningarXivPaperNabila Tasnim, Haoran Liu et al.Jul 22
028WaveformQA: Benchmarking LLM Temporal Reasoning on Digital WaveformsarXivPaperYichuan Liu, Daniel G Cummings et al.Jul 22
029BaseRT: Advancing Best-in-Class LLM Inference with Apple M5 Neural AcceleratorsarXivPaperFabian Waschkowski, Prabod Rathnayaka et al.Jul 21
030From Bit-Position Sensitivity to Unequal Error Protection for DNN Inference MemoryarXivPaperMuhammad Husnain Mubarik, Karthik Mohan Kumar et al.Jul 21
031BRIM: Workload-Balanced Dual-Sided Bit-Serial Sparse Inference AcceleratorarXivPaperVarun Manjunath, Ruokai Yin et al.Jul 20
032Can AI Agents Really Complete RTL-to-GDS? Lessons from Benchmarking Tool-Interactive EDA WorkflowsarXivPaperJinyuan Deng, Zhengrui Chen et al.Jul 20
033D-NOVA: In-Storage Retrieval Accelerator via Dual-Bound 3D NAND-Optimized Similarity Search with Vector AdaptationarXivPaperChang Eun Song, Sumukh Pinge et al.Jul 20
034Empowering On-Device Model Adaptation with an Edge AI Inference AcceleratorarXivPaperM. Piechocki, Alessandro Capotondi et al.Jul 20
035Hardware Mechanisms to Dynamically Throttle AI PerformancearXivPaperHaiyue Ma, Lauren Malek et al.Jul 20
036Isolation Failure From Shared Storage: Characterizing and Exploiting Page-Cache SCA Leakage Across Containers and VMsarXivPaperAlon Abudraham, Xingyu Chen et al.Jul 20
037Opto-ViT-v2: Noise-Resilient On-Chip Fine-Tuning for Photonic Near-Sensor Vision Transformer AcceleratorsarXivPaperXuming Chen et al.Jul 20
038PIP-NTT: Towards a Scalable Memory-Parallelized Accelerator for Iterative NTT in PQCarXivPaperMalik Imran, A. Khalid et al.Jul 20
039PRISM: Sensitivity-Aware PolynoMial PRuning for EffIcient Neural Network EncryptionarXivPaperSahaj Majavdia, Mahdi TaheriJul 20
040ThAME: 3D Memory-Enabled Heterogeneous Accelerator for LLM Mixture of ExpertsarXivPaperPratyush Dhingra, Pramit Kumar Pal et al.Jul 19
041ThRIve: Thermally Robust CNN Inference via Low-Rank Adaptation in Heterogeneous PIM ArchitecturesarXivPaperVibhanshu Sharma, Pratyush Dhingra et al.Jul 19
042SABLE: Minimalist Instruction-Level Authenticated Encryption for Constrained Confidential ComputingarXivPaperHamid Noori, Carlton ShepherdJul 18
043CoG-Guided Weight Correction for Fault-Tolerant Deep Neural NetworksarXivPaperBahram Parchekani, Samira Nazari et al.Jul 17
044Mitigating Compiler Fusion-Induced Power Bursts in Mobile NPU Inference as the Battery DepletesarXivPaperRyoga Yuzawa, Masayoshi TomizukaJul 17
045RTL-Sequencer: Towards Scalable RTL Timing Prediction with the Sequence-based ParadigmarXivPaperZiyan Guo, Wenji Fang et al.Jul 17
046Vogls: a Fast Interactive Full-timing Simulator for Pre-silicon Power Side-Channel AnalysisarXivPaperGijs Burghoorn, I. Buhan et al.Jul 17
047ExaGEMM: Exploration Framework for CPU-Driven ML Inference via Associative In-Register Computing for Low-Bit GEMMarXivPaperHyunwoo Oh, Suyeon Jang et al.Jul 16
048Lazy Arithmetic using Systolic Arrays for Closing the Verification Gap on Embedded SystemsarXivPaperTaisa Kushner, Ryan McCleeary et al.Jul 16
049NIFA: Nonlinear IMC enhanced FPGA for efficient ML inferencearXivPaperJiajun Hu, R. Sunketa et al.Jul 16
050PolyQ: Codesigning End-to-End Quantization Framework for Scalable Edge CPU LLM InferencearXivPaperHyunwoo Oh, Suyeon Jang et al.Jul 16
051Toward Energy-Efficient and Low-Power Arrhythmia Detection for Wearable DevicesarXivPaperF. Bulten, Yawar Rasheed et al.Jul 16
052CODA: How to Mitigate ColumnDisturb for (Almost) Free?arXivPaperMoinuddin K. QureshiJul 15
053Kaleido: Algorithm-Hardware Co-Design for Video Diffusion Transformers by Exploiting Latent Space CorrelationsarXivPaperWenxuan Miao, Haosong Liu et al.Jul 15
054Lighthouse RL: Sample-Efficient Circuit Optimization via Strategic Reset PointsarXivPaperMustafa Emre Gursoy, Stefan Uhlich et al.Jul 15
055Towards Reliable AI-Assisted Analog Design: Template-Constrained LLM Agents for SAR ADC GenerationarXivPaperD. Kochar, Hae-Seung Lee et al.Jul 15
056A 32-channel event-based bio-signal analog front-end with adaptive delta and pulse frequency encodingarXivPaperNarayanan Shyam et al.Jul 14
057Emulated Integrity Replica: Enabling Self-Healing on FPGA SoCs via Hierarchical TwinsarXivPaperArsalan Ali Malik, Ali Suvizi et al.Jul 14
058Full-Pipeline Inference Optimization for MiMo-V2.5 Series: Pushing Hybrid SWA Efficiency to the LimitarXivPaperXiaomin Liu, Aoxin Ma et al.Jul 14
059Can LLMs Perform Deep Technical Comprehension of Computer Architecture Papers?arXivPaperNishant Aggarwal, A. Dubal et al.Jul 13
060HiFi-LLP: High-Fidelity, Low-Cost Latency Predictors with Confidence for Robust HW-NASarXivPaperShambhavi Balamuthu Sampath, Behzad Shomali et al.Jul 13

Showing 60 of 3,552 documents · scroll for more