Skip to main content

Showing 1–50 of 144 results for author: Zeng, K

.
  1. arXiv:2406.20083  [pdf, other

    cs.RO cs.CV

    PoliFormer: Scaling On-Policy RL with Transformers Results in Masterful Navigators

    Authors: Kuo-Hao Zeng, Zichen Zhang, Kiana Ehsani, Rose Hendrix, Jordi Salvador, Alvaro Herrasti, Ross Girshick, Aniruddha Kembhavi, Luca Weihs

    Abstract: We present PoliFormer (Policy Transformer), an RGB-only indoor navigation agent trained end-to-end with reinforcement learning at scale that generalizes to the real-world without adaptation despite being trained purely in simulation. PoliFormer uses a foundational vision transformer encoder with a causal transformer decoder enabling long-term memory and reasoning. It is trained for hundreds of mil… ▽ More

    Submitted 28 June, 2024; originally announced June 2024.

  2. arXiv:2406.02624  [pdf, other

    cs.CR cs.SE

    Take a Step Further: Understanding Page Spray in Linux Kernel Exploitation

    Authors: Ziyi Guo, Dang K Le, Zhenpeng Lin, Kyle Zeng, Ruoyu Wang, Tiffany Bao, Yan Shoshitaishvili, Adam Doupé, Xinyu Xing

    Abstract: Recently, a novel method known as Page Spray emerges, focusing on page-level exploitation for kernel vulnerabilities. Despite the advantages it offers in terms of exploitability, stability, and compatibility, comprehensive research on Page Spray remains scarce. Questions regarding its root causes, exploitation model, comparative benefits over other exploitation techniques, and possible mitigation… ▽ More

    Submitted 6 June, 2024; v1 submitted 3 June, 2024; originally announced June 2024.

  3. arXiv:2405.17233  [pdf, other

    cs.LG

    CLAQ: Pushing the Limits of Low-Bit Post-Training Quantization for LLMs

    Authors: Haoyu Wang, Bei Liu, Hang Shao, Bo Xiao, Ke Zeng, Guanglu Wan, Yanmin Qian

    Abstract: Parameter quantization for Large Language Models (LLMs) has attracted increasing attentions recently in reducing memory costs and improving computational efficiency. Early approaches have been widely adopted. However, the existing methods suffer from poor performance in low-bit (such as 2 to 3 bits) scenarios. In this paper, we present a novel and effective Column-Level Adaptive weight Quantizatio… ▽ More

    Submitted 2 June, 2024; v1 submitted 27 May, 2024; originally announced May 2024.

  4. arXiv:2405.13596  [pdf, other

    astro-ph.HE astro-ph.SR

    SN 2023zaw: the low-energy explosion of an ultra-stripped star, with non-radioactive heating

    Authors: Thomas Moore, James Gillanders, Matt Nicholl, Mark Huber, Stephen Smartt, Shubham Srivastav, Heloise Stevance, Ting-Wan Chen, Kenneth Chambers, Joseph Anderson, Michael Fulton, Samantha Oates, Charlotte Angus, Giuliano Pignata, Nicolas Erasmus, Hua Gao, Joanna Bulger, Chien-Cheng Lin, Thomas Lowe, Eugene Magnier, Paloma Minguez, Chow-Choong Ngeow, Xinyue Sheng, Stuart A. Sim, Ken Smith , et al. (4 additional authors not shown)

    Abstract: Most stripped envelope supernova progenitors are formed through binary interaction, losing hydrogen and/or helium from their outer layers. An emerging class of supernovae with the highest degree of envelope-strip** are thought to be the product of strip** by a NS companion. However, relatively few examples are known and the outcomes of such systems can be diverse and are poorly understood at p… ▽ More

    Submitted 22 May, 2024; originally announced May 2024.

  5. Swipe2Pair: Secure and Fast In-Band Wireless Device Pairing

    Authors: Yaqi He, Kai Zeng, Long Jiao, Brian L. Mark, Khaled N. Khasawneh

    Abstract: Wireless device pairing is a critical security mechanism to bootstrap the secure communication between two devices without a pre-shared secret. It has been widely used in many Internet of Things (IoT) applications, such as smart-home and smart-health. Most existing device pairing mechanisms are based on out-of-band channels, e.g., extra sensors or hardware, to validate the proximity of pairing dev… ▽ More

    Submitted 5 May, 2024; originally announced May 2024.

  6. arXiv:2404.12794  [pdf, other

    cs.CV cs.MM cs.RO eess.IV

    MambaMOS: LiDAR-based 3D Moving Object Segmentation with Motion-aware State Space Model

    Authors: Kang Zeng, Hao Shi, Jiacheng Lin, Siyu Li, **tao Cheng, Kaiwei Wang, Zhiyong Li, Kailun Yang

    Abstract: LiDAR-based Moving Object Segmentation (MOS) aims to locate and segment moving objects in point clouds of the current scan using motion information from previous scans. Despite the promising results achieved by previous MOS methods, several key issues, such as the weak coupling of temporal and spatial information, still need further study. In this paper, we propose a novel LiDAR-based 3D Moving Ob… ▽ More

    Submitted 19 April, 2024; originally announced April 2024.

    Comments: The source code will be made publicly available at https://github.com/Terminal-K/MambaMOS

  7. arXiv:2404.12242  [pdf, other

    cs.CL

    CMNEE: A Large-Scale Document-Level Event Extraction Dataset based on Open-Source Chinese Military News

    Authors: Mengna Zhu, Zijie Xu, Kaisheng Zeng, Kaiming Xiao, Mao Wang, Wenjun Ke, Hongbin Huang

    Abstract: Extracting structured event knowledge, including event triggers and corresponding arguments, from military texts is fundamental to many applications, such as intelligence analysis and decision assistance. However, event extraction in the military field faces the data scarcity problem, which impedes the research of event extraction models in this domain. To alleviate this problem, we propose CMNEE,… ▽ More

    Submitted 18 April, 2024; originally announced April 2024.

    Comments: 13 pages, 7 figures, accepted to LREC-COLING 2024

  8. arXiv:2404.04956  [pdf, other

    cs.CV cs.CR

    Gaussian Shading: Provable Performance-Lossless Image Watermarking for Diffusion Models

    Authors: Zi** Yang, Kai Zeng, Kejiang Chen, Han Fang, Weiming Zhang, Nenghai Yu

    Abstract: Ethical concerns surrounding copyright protection and inappropriate content generation pose challenges for the practical implementation of diffusion models. One effective solution involves watermarking the generated images. However, existing methods often compromise the model performance or require additional training, which is undesirable for operators and users. To address this issue, we propose… ▽ More

    Submitted 6 May, 2024; v1 submitted 7 April, 2024; originally announced April 2024.

    Comments: 17 pages, 11 figures, accepted by CVPR 2024

  9. arXiv:2404.00044  [pdf, other

    physics.chem-ph cs.AI cs.LG q-bio.QM

    UAlign: Pushing the Limit of Template-free Retrosynthesis Prediction with Unsupervised SMILES Alignment

    Authors: Kaipeng Zeng, Bo yang, Xin Zhao, Yu Zhang, Fan Nie, Xiaokang Yang, Yaohui **, Yanyan Xu

    Abstract: Motivation: Retrosynthesis planning poses a formidable challenge in the organic chemical industry. Single-step retrosynthesis prediction, a crucial step in the planning process, has witnessed a surge in interest in recent years due to advancements in AI for science. Various deep learning-based methods have been proposed for this task in recent years, incorporating diverse levels of additional chem… ▽ More

    Submitted 19 April, 2024; v1 submitted 24 March, 2024; originally announced April 2024.

  10. arXiv:2403.20188  [pdf, other

    cs.NI cs.AI cs.LG

    Distributed Swarm Learning for Edge Internet of Things

    Authors: Yue Wang, Zhi Tian, FXin Fan, Zhipeng Cai, Cameron Nowzari, Kai Zeng

    Abstract: The rapid growth of Internet of Things (IoT) has led to the widespread deployment of smart IoT devices at wireless edge for collaborative machine learning tasks, ushering in a new era of edge learning. With a huge number of hardware-constrained IoT devices operating in resource-limited wireless networks, edge learning encounters substantial challenges, including communication and computation bottl… ▽ More

    Submitted 29 March, 2024; originally announced March 2024.

    Comments: arXiv admin note: substantial text overlap with arXiv:2210.16705

  11. arXiv:2403.17524  [pdf, other

    cs.CR cs.CL

    Provably Secure Disambiguating Neural Linguistic Steganography

    Authors: Yuang Qi, Kejiang Chen, Kai Zeng, Weiming Zhang, Nenghai Yu

    Abstract: Recent research in provably secure neural linguistic steganography has overlooked a crucial aspect: the sender must detokenize stegotexts to avoid raising suspicion from the eavesdropper. The segmentation ambiguity problem, which arises when using language models based on subwords, leads to occasional decoding failures in all neural language steganography implementations based on these models. Cur… ▽ More

    Submitted 26 March, 2024; originally announced March 2024.

  12. arXiv:2403.03201  [pdf, other

    cond-mat.str-el

    Site Symmetry and Multiorbital Flat Bands on Kagome and Pyrochlore Lattices

    Authors: Keyu Zeng, Ziqiang Wang

    Abstract: Flat bands in electronic band structures are intriguing platforms for strong correlation and topological physics, primarily due to the suppressed kinetic energy of electrons. Various methods have been developed to create flat bands, utilizing lattice geometry or finely tuned parameters. Despite this, the investigation of orbital symmetry in multiorbital materials is a relatively new area of focus.… ▽ More

    Submitted 14 March, 2024; v1 submitted 5 March, 2024; originally announced March 2024.

  13. arXiv:2402.18243  [pdf, other

    cs.CL

    Learning or Self-aligning? Rethinking Instruction Fine-tuning

    Authors: Mengjie Ren, Boxi Cao, Hongyu Lin, Cao Liu, Xianpei Han, Ke Zeng, Guanglu Wan, Xunliang Cai, Le Sun

    Abstract: Instruction Fine-tuning~(IFT) is a critical phase in building large language models~(LLMs). Previous works mainly focus on the IFT's role in the transfer of behavioral norms and the learning of additional world knowledge. However, the understanding of the underlying mechanisms of IFT remains significantly limited. In this paper, we design a knowledge intervention framework to decouple the potentia… ▽ More

    Submitted 2 March, 2024; v1 submitted 28 February, 2024; originally announced February 2024.

  14. arXiv:2402.16578  [pdf, other

    cs.CL cs.LG

    Multi-Bit Distortion-Free Watermarking for Large Language Models

    Authors: Massieh Kordi Boroujeny, Ya Jiang, Kai Zeng, Brian Mark

    Abstract: Methods for watermarking large language models have been proposed that distinguish AI-generated text from human-generated text by slightly altering the model output distribution, but they also distort the quality of the text, exposing the watermark to adversarial detection. More recently, distortion-free watermarking methods were proposed that require a secret key to detect the watermark. The prio… ▽ More

    Submitted 26 February, 2024; originally announced February 2024.

  15. arXiv:2402.13093  [pdf, other

    cs.CL cs.AI

    Event-level Knowledge Editing

    Authors: Hao Peng, Xiaozhi Wang, Chunyang Li, Kaisheng Zeng, Jiangshan Duo, Yixin Cao, Lei Hou, Juanzi Li

    Abstract: Knowledge editing aims at updating knowledge of large language models (LLMs) to prevent them from becoming outdated. Existing work edits LLMs at the level of factual knowledge triplets. However, natural knowledge updates in the real world come from the occurrences of new events rather than direct changes in factual triplets. In this paper, we propose a new task setting: event-level knowledge editi… ▽ More

    Submitted 21 April, 2024; v1 submitted 20 February, 2024; originally announced February 2024.

    Comments: 18 pages, 2 figures

  16. arXiv:2401.17023  [pdf, other

    cs.CV

    MF-MOS: A Motion-Focused Model for Moving Object Segmentation

    Authors: **tao Cheng, Kang Zeng, Zhuoxu Huang, Xiaoyu Tang, ** Wu, Chengxi Zhang, Xieyuanli Chen, Rui Fan

    Abstract: Moving object segmentation (MOS) provides a reliable solution for detecting traffic participants and thus is of great interest in the autonomous driving field. Dynamic capture is always critical in the MOS problem. Previous methods capture motion features from the range images directly. Differently, we argue that the residual maps provide greater potential for motion information, while range image… ▽ More

    Submitted 30 January, 2024; originally announced January 2024.

    Comments: Accepted by ICRA2024

  17. arXiv:2401.09500  [pdf, other

    q-bio.NC cs.LG cs.NE

    MorphGrower: A Synchronized Layer-by-layer Growing Approach for Plausible Neuronal Morphology Generation

    Authors: Nianzu Yang, Kaipeng Zeng, Haotian Lu, Yexin Wu, Zexin Yuan, Danni Chen, Shengdian Jiang, Jiaxiang Wu, Yimin Wang, Junchi Yan

    Abstract: Neuronal morphology is essential for studying brain functioning and understanding neurodegenerative disorders. As acquiring real-world morphology data is expensive, computational approaches for morphology generation have been studied. Traditional methods heavily rely on expert-set rules and parameter tuning, making it difficult to generalize across different types of morphologies. Recently, MorphV… ▽ More

    Submitted 27 May, 2024; v1 submitted 17 January, 2024; originally announced January 2024.

  18. arXiv:2401.07770  [pdf, other

    cs.CV

    Seeing the Unseen: Visual Common Sense for Semantic Placement

    Authors: Ram Ramrakhya, Aniruddha Kembhavi, Dhruv Batra, Zsolt Kira, Kuo-Hao Zeng, Luca Weihs

    Abstract: Computer vision tasks typically involve describing what is present in an image (e.g. classification, detection, segmentation, and captioning). We study a visual common sense task that requires understanding what is not present. Specifically, given an image (e.g. of a living room) and name of an object ("cushion"), a vision system is asked to predict semantically-meaningful regions (masks or boundi… ▽ More

    Submitted 15 January, 2024; originally announced January 2024.

  19. arXiv:2312.02976  [pdf, other

    cs.RO cs.AI cs.CV

    Imitating Shortest Paths in Simulation Enables Effective Navigation and Manipulation in the Real World

    Authors: Kiana Ehsani, Tanmay Gupta, Rose Hendrix, Jordi Salvador, Luca Weihs, Kuo-Hao Zeng, Kunal Pratap Singh, Ye** Kim, Winson Han, Alvaro Herrasti, Ranjay Krishna, Dustin Schwenk, Eli VanderBilt, Aniruddha Kembhavi

    Abstract: Reinforcement learning (RL) with dense rewards and imitation learning (IL) with human-generated trajectories are the most widely used approaches for training modern embodied agents. RL requires extensive reward sha** and auxiliary losses and is often too slow and ineffective for long-horizon tasks. While IL with human supervision is effective, collecting human trajectories at scale is extremely… ▽ More

    Submitted 5 December, 2023; originally announced December 2023.

    Comments: First six authors contributed equally. Project page: https://spoc-robot.github.io/

  20. arXiv:2311.09105  [pdf, other

    cs.CL

    MAVEN-Arg: Completing the Puzzle of All-in-One Event Understanding Dataset with Event Argument Annotation

    Authors: Xiaozhi Wang, Hao Peng, Yong Guan, Kaisheng Zeng, Jianhui Chen, Lei Hou, Xu Han, Yankai Lin, Zhiyuan Liu, Ruobing Xie, Jie Zhou, Juanzi Li

    Abstract: Understanding events in texts is a core objective of natural language understanding, which requires detecting event occurrences, extracting event arguments, and analyzing inter-event relationships. However, due to the annotation challenges brought by task complexity, a large-scale dataset covering the full process of event understanding has long been absent. In this paper, we introduce MAVEN-Arg,… ▽ More

    Submitted 18 June, 2024; v1 submitted 15 November, 2023; originally announced November 2023.

    Comments: Accepted at ACL 2024. Camera-ready version

  21. arXiv:2311.08993  [pdf, other

    cs.CL cs.AI

    When does In-context Learning Fall Short and Why? A Study on Specification-Heavy Tasks

    Authors: Hao Peng, Xiaozhi Wang, Jianhui Chen, Weikai Li, Yunjia Qi, Zimu Wang, Zhili Wu, Kaisheng Zeng, Bin Xu, Lei Hou, Juanzi Li

    Abstract: In-context learning (ICL) has become the default method for using large language models (LLMs), making the exploration of its limitations and understanding the underlying causes crucial. In this paper, we find that ICL falls short of handling specification-heavy tasks, which are tasks with complicated and extensive task specifications, requiring several hours for ordinary humans to master, such as… ▽ More

    Submitted 15 November, 2023; originally announced November 2023.

    Comments: Under review

  22. arXiv:2311.05168  [pdf, other

    cs.CV cs.AI

    FireMatch: A Semi-Supervised Video Fire Detection Network Based on Consistency and Distribution Alignment

    Authors: Qinghua Lin, Zuoyong Li, Kun Zeng, Haoyi Fan, Wei Li, Xiaoguang Zhou

    Abstract: Deep learning techniques have greatly enhanced the performance of fire detection in videos. However, video-based fire detection models heavily rely on labeled data, and the process of data labeling is particularly costly and time-consuming, especially when dealing with videos. Considering the limited quantity of labeled video data, we propose a semi-supervised fire detection model called FireMatch… ▽ More

    Submitted 9 November, 2023; originally announced November 2023.

  23. arXiv:2311.04193  [pdf, other

    cs.CV cs.AI

    Selective Visual Representations Improve Convergence and Generalization for Embodied AI

    Authors: Ainaz Eftekhar, Kuo-Hao Zeng, Jiafei Duan, Ali Farhadi, Ani Kembhavi, Ranjay Krishna

    Abstract: Embodied AI models often employ off the shelf vision backbones like CLIP to encode their visual observations. Although such general purpose representations encode rich syntactic and semantic information about the scene, much of this information is often irrelevant to the specific task at hand. This introduces noise within the learning process and distracts the agent's focus from task-relevant visu… ▽ More

    Submitted 9 March, 2024; v1 submitted 7 November, 2023; originally announced November 2023.

    Comments: See project website: https://embodied-codebook.github.io

  24. arXiv:2310.10590  [pdf, other

    cs.CL

    Mastering the Task of Open Information Extraction with Large Language Models and Consistent Reasoning Environment

    Authors: Ji Qi, Kaixuan Ji, Xiaozhi Wang, Jifan Yu, Kaisheng Zeng, Lei Hou, Juanzi Li, Bin Xu

    Abstract: Open Information Extraction (OIE) aims to extract objective structured knowledge from natural texts, which has attracted growing attention to build dedicated models with human experience. As the large language models (LLMs) have exhibited remarkable in-context learning capabilities, a question arises as to whether the task of OIE can be effectively tackled with this paradigm? In this paper, we exp… ▽ More

    Submitted 16 October, 2023; originally announced October 2023.

  25. arXiv:2310.09499  [pdf, other

    cs.CL cs.AI

    One-Shot Sensitivity-Aware Mixed Sparsity Pruning for Large Language Models

    Authors: Hang Shao, Bei Liu, Bo Xiao, Ke Zeng, Guanglu Wan, Yanmin Qian

    Abstract: Various Large Language Models~(LLMs) from the Generative Pretrained Transformer(GPT) family have achieved outstanding performances in a wide range of text generation tasks. However, the enormous model sizes have hindered their practical use in real-world applications due to high inference latency. Therefore, improving the efficiencies of LLMs through quantization, pruning, and other means has been… ▽ More

    Submitted 23 April, 2024; v1 submitted 14 October, 2023; originally announced October 2023.

    Comments: Accepted to ICASSP2024

  26. arXiv:2310.08864  [pdf, other

    cs.RO

    Open X-Embodiment: Robotic Learning Datasets and RT-X Models

    Authors: Open X-Embodiment Collaboration, Abby O'Neill, Abdul Rehman, Abhinav Gupta, Abhiram Maddukuri, Abhishek Gupta, Abhishek Padalkar, Abraham Lee, Acorn Pooley, Agrim Gupta, Ajay Mandlekar, A**kya Jain, Albert Tung, Alex Bewley, Alex Herzog, Alex Irpan, Alexander Khazatsky, Anant Rai, Anchit Gupta, Andrew Wang, Andrey Kolobov, Anikait Singh, Animesh Garg, Aniruddha Kembhavi, Annie Xie , et al. (267 additional authors not shown)

    Abstract: Large, high-capacity models trained on diverse datasets have shown remarkable successes on efficiently tackling downstream applications. In domains from NLP to Computer Vision, this has led to a consolidation of pretrained models, with general pretrained backbones serving as a starting point for many applications. Can such a consolidation happen in robotics? Conventionally, robotic learning method… ▽ More

    Submitted 1 June, 2024; v1 submitted 13 October, 2023; originally announced October 2023.

    Comments: Project website: https://robotics-transformer-x.github.io

  27. arXiv:2310.08027  [pdf, other

    cs.CL cs.CV

    Exploring Large Language Models for Multi-Modal Out-of-Distribution Detection

    Authors: Yi Dai, Hao Lang, Kaisheng Zeng, Fei Huang, Yongbin Li

    Abstract: Out-of-distribution (OOD) detection is essential for reliable and trustworthy machine learning. Recent multi-modal OOD detection leverages textual information from in-distribution (ID) class names for visual OOD detection, yet it currently neglects the rich contextual information of ID classes. Large language models (LLMs) encode a wealth of world knowledge and can be prompted to generate descript… ▽ More

    Submitted 12 October, 2023; originally announced October 2023.

    Comments: EMNLP2023 Findings Long Paper

  28. arXiv:2310.06417  [pdf, other

    cs.LG cs.AI

    Advective Diffusion Transformers for Topological Generalization in Graph Learning

    Authors: Qitian Wu, Chenxiao Yang, Kaipeng Zeng, Fan Nie, Michael Bronstein, Junchi Yan

    Abstract: Graph diffusion equations are intimately related to graph neural networks (GNNs) and have recently attracted attention as a principled framework for analyzing GNN dynamics, formalizing their expressive power, and justifying architectural choices. One key open questions in graph learning is the generalization capabilities of GNNs. A major limitation of current approaches hinges on the assumption th… ▽ More

    Submitted 10 October, 2023; originally announced October 2023.

    Comments: 39 pages

  29. arXiv:2310.00597  [pdf, other

    cs.CL

    A Task-oriented Dialog Model with Task-progressive and Policy-aware Pre-training

    Authors: Lucen Zhong, Hengtong Lu, Caixia Yuan, Xiaojie Wang, Jiashen Sun, Ke Zeng, Guanglu Wan

    Abstract: Pre-trained conversation models (PCMs) have achieved promising progress in recent years. However, existing PCMs for Task-oriented dialog (TOD) are insufficient for capturing the sequential nature of the TOD-related tasks, as well as for learning dialog policy information. To alleviate these problems, this paper proposes a task-progressive PCM with two policy-aware pre-training tasks. The model is… ▽ More

    Submitted 1 October, 2023; originally announced October 2023.

    Comments: Accepted at NLPCC 2023

  30. arXiv:2309.14258  [pdf, other

    cs.CL cs.AI

    OmniEvent: A Comprehensive, Fair, and Easy-to-Use Toolkit for Event Understanding

    Authors: Hao Peng, Xiaozhi Wang, Feng Yao, Zimu Wang, Chuzhao Zhu, Kaisheng Zeng, Lei Hou, Juanzi Li

    Abstract: Event understanding aims at understanding the content and relationship of events within texts, which covers multiple complicated information extraction tasks: event detection, event argument extraction, and event relation extraction. To facilitate related research and application, we present an event understanding toolkit OmniEvent, which features three desiderata: (1) Comprehensive. OmniEvent sup… ▽ More

    Submitted 25 September, 2023; originally announced September 2023.

  31. arXiv:2308.14603  [pdf

    physics.optics

    Recovering lossless propagation of polaritons with synthesized complex frequency excitation

    Authors: Fuxin Guan, Xiangdong Guo, Shu Zhang, Kebo Zeng, Yue Hu, Chenchen Wu, Shaobo Zhou, Yuanjiang Xiang, Xiaoxia Yang, Qing Dai, Shuang Zhang

    Abstract: Surface plasmon polaritons and phonon polaritons offer a means of surpassing the diffraction limit of conventional optics and facilitate efficient energy storage, local field enhancement, high sensitivities, benefitting from their subwavelength confinement of light. Unfortunately, losses severely limit the propagation decay length, thus restricting the practical use of polaritons. While optimizing… ▽ More

    Submitted 18 September, 2023; v1 submitted 28 August, 2023; originally announced August 2023.

    Comments: 20 pages, 4 figures

  32. arXiv:2308.13884  [pdf, ps, other

    cs.NI

    Location Privacy and Spectrum Efficiency Enhancement in Spectrum Sharing Systems

    Authors: Long Jiao, Yao Ge, Kai Zeng, B. C. Hilburn

    Abstract: In this work, we investigate the benefits of secondary user (SU) network beamforming on improving primary user (PU) location privacy in spectrum sharing systems, where the beamformer in the SU network is designed to suppress the aggregate interference to improve the location privacy of PUs. We consider two problems: improving SU network communication throughput subject to the specified PU location… ▽ More

    Submitted 26 August, 2023; originally announced August 2023.

  33. Optically levitated gyroscopes with a MHz rotating micro-rotor

    Authors: Kai Zeng, Xiangming Xu, Yulie Wu, Xuezhong Wu, Dingbang Xiao

    Abstract: The optically levitated particles have been driven to rotate at an ultra-high speed of GHz, and the gyroscopic application of these levitated particles to measure angular motion have long been explored. However, this gyroscope has not been proven either theoretically or experimentally. Here, a rotor gyroscope based on optically levitated high-speed rotating particles is proposed. In vacuum, an ell… ▽ More

    Submitted 17 August, 2023; originally announced August 2023.

    Journal ref: Microsystems & Nanoengineering, 10 78 (2024)

  34. arXiv:2307.10545  [pdf, ps, other

    math.QA hep-th math.AT

    Loday-Quillen-Tsygan theorem on Quivers

    Authors: Keyou Zeng

    Abstract: The well-known Loday-Quillen-Tsygan theorem calculates the Lie algebra homology of the infinite general linear Lie algebra $\mathfrak{gl}(A)$ over an unital associative algebra $A$. We generalize the Loday-Quillen-Tsygan theorem to an infinite Lie algebra associated with a (framed) quiver, where we assign to each vertex $v$ an infinite general linear Lie algebra $\mathfrak{gl}(A_v)$, to each edge… ▽ More

    Submitted 19 July, 2023; originally announced July 2023.

  35. arXiv:2307.10537  [pdf, other

    quant-ph physics.app-ph

    Wide-band Unambiguous Quantum Sensing via Geodesic Evolution

    Authors: Ke Zeng, Xiaohui Yu, Martin B. Plenio, Zhen-Yu Wang

    Abstract: We present a quantum sensing technique that utilizes a sequence of $π$ pulses to cyclically drive the qubit dynamics along a geodesic path of adiabatic evolution. This approach effectively suppresses the effects of both decoherence noise and control errors while simultaneously removing unwanted resonance terms, such as higher harmonics and spurious responses commonly encountered in dynamical decou… ▽ More

    Submitted 19 July, 2023; originally announced July 2023.

  36. arXiv:2307.09117  [pdf

    physics.optics physics.app-ph

    Synthesized complex-frequency excitation for ultrasensitive molecular sensing

    Authors: Kebo Zeng, Chenchen Wu, Xiangdong Guo, Fuxin Guan, Yu Duan, Lauren L Zhang, Xiaoxia Yang, Na Liu, Qing Dai, Shuang Zhang

    Abstract: Detecting trace molecules remains a significant challenge. Surface-enhanced infrared absorption (SEIRA) based on plasmonic nanostructures, particularly graphene, has emerged as a promising approach to enhance sensing sensitivity. While graphene-based SEIRA offers advantages such as ultrahigh sensitivity and active tunability, intrinsic molecular dam** weakens the interaction between vibrational… ▽ More

    Submitted 18 July, 2023; originally announced July 2023.

    Comments: 21 pages, 4 figures

  37. arXiv:2306.09296  [pdf, other

    cs.CL

    KoLA: Carefully Benchmarking World Knowledge of Large Language Models

    Authors: Jifan Yu, Xiaozhi Wang, Shangqing Tu, Shulin Cao, Daniel Zhang-Li, Xin Lv, Hao Peng, Zijun Yao, Xiaohan Zhang, Hanming Li, Chunyang Li, Zheyuan Zhang, Yushi Bai, Yantao Liu, Amy Xin, Nianyi Lin, Kaifeng Yun, Linlu Gong, Jianhui Chen, Zhili Wu, Yunjia Qi, Weikai Li, Yong Guan, Kaisheng Zeng, Ji Qi , et al. (10 additional authors not shown)

    Abstract: The unprecedented performance of large language models (LLMs) necessitates improvements in evaluations. Rather than merely exploring the breadth of LLM abilities, we believe meticulous and thoughtful designs are essential to thorough, unbiased, and applicable evaluations. Given the importance of world knowledge to LLMs, we construct a Knowledge-oriented LLM Assessment benchmark (KoLA), in which we… ▽ More

    Submitted 30 June, 2024; v1 submitted 15 June, 2023; originally announced June 2023.

    Comments: Accepted by ICLR 2024

  38. arXiv:2306.06918  [pdf, other

    cs.CL cs.AI

    The Devil is in the Details: On the Pitfalls of Event Extraction Evaluation

    Authors: Hao Peng, Xiaozhi Wang, Feng Yao, Kaisheng Zeng, Lei Hou, Juanzi Li, Zhiyuan Liu, Weixing Shen

    Abstract: Event extraction (EE) is a crucial task aiming at extracting events from texts, which includes two subtasks: event detection (ED) and event argument extraction (EAE). In this paper, we check the reliability of EE evaluations and identify three major pitfalls: (1) The data preprocessing discrepancy makes the evaluation results on the same dataset not directly comparable, but the data preprocessing… ▽ More

    Submitted 15 June, 2023; v1 submitted 12 June, 2023; originally announced June 2023.

    Comments: Accepted at Findings of ACL 2023

  39. arXiv:2306.04181  [pdf, other

    cs.CL cs.LG

    Benchmarking Foundation Models with Language-Model-as-an-Examiner

    Authors: Yushi Bai, Jiahao Ying, Yixin Cao, Xin Lv, Yuze He, Xiaozhi Wang, Jifan Yu, Kaisheng Zeng, Yijia Xiao, Haozhe Lyu, Jiayin Zhang, Juanzi Li, Lei Hou

    Abstract: Numerous benchmarks have been established to assess the performance of foundation models on open-ended question answering, which serves as a comprehensive test of a model's ability to understand and generate language in a manner similar to humans. Most of these works focus on proposing new datasets, however, we see two main issues within previous benchmarking pipelines, namely testing leakage and… ▽ More

    Submitted 4 November, 2023; v1 submitted 7 June, 2023; originally announced June 2023.

    Comments: NeurIPS 2023 Datasets and Benchmarks

  40. arXiv:2305.13981  [pdf, other

    cs.CL cs.AI

    Preserving Knowledge Invariance: Rethinking Robustness Evaluation of Open Information Extraction

    Authors: Ji Qi, Chuchun Zhang, Xiaozhi Wang, Kaisheng Zeng, Jifan Yu, **xin Liu, Jiuding Sun, Yuxiang Chen, Lei Hou, Juanzi Li, Bin Xu

    Abstract: The robustness to distribution changes ensures that NLP models can be successfully applied in the realistic world, especially for information extraction tasks. However, most prior evaluation benchmarks have been devoted to validating pairwise matching correctness, ignoring the crucial measurement of robustness. In this paper, we present the first benchmark that simulates the evaluation of open inf… ▽ More

    Submitted 24 October, 2023; v1 submitted 23 May, 2023; originally announced May 2023.

    Comments: Accepted by EMNLP 2023 Main Conference

  41. arXiv:2305.10863  [pdf, other

    cs.DC cs.AI cs.LG cs.OS

    Quiver: Supporting GPUs for Low-Latency, High-Throughput GNN Serving with Workload Awareness

    Authors: Zeyuan Tan, Xiulong Yuan, Congjie He, Man-Kit Sit, Guo Li, Xiaoze Liu, Baole Ai, Kai Zeng, Peter Pietzuch, Luo Mai

    Abstract: Systems for serving inference requests on graph neural networks (GNN) must combine low latency with high throughout, but they face irregular computation due to skew in the number of sampled graph nodes and aggregated GNN features. This makes it challenging to exploit GPUs effectively: using GPUs to sample only a few graph nodes yields lower performance than CPU-based sampling; and aggregating many… ▽ More

    Submitted 18 May, 2023; originally announced May 2023.

  42. arXiv:2305.01090  [pdf, ps, other

    cs.LG nlin.CD

    Autoencoders for discovering manifold dimension and coordinates in data from complex dynamical systems

    Authors: Kevin Zeng, Carlos E. Pérez De Jesús, Andrew J. Fox, Michael D. Graham

    Abstract: While many phenomena in physics and engineering are formally high-dimensional, their long-time dynamics often live on a lower-dimensional manifold. The present work introduces an autoencoder framework that combines implicit regularization with internal linear layers and $L_2$ regularization (weight decay) to automatically estimate the underlying dimensionality of a data set, produce an orthogonal… ▽ More

    Submitted 6 December, 2023; v1 submitted 1 May, 2023; originally announced May 2023.

  43. arXiv:2304.12289  [pdf, other

    cs.CV cs.AI cs.RO

    Moving Forward by Moving Backward: Embedding Action Impact over Action Semantics

    Authors: Kuo-Hao Zeng, Luca Weihs, Roozbeh Mottaghi, Ali Farhadi

    Abstract: A common assumption when training embodied agents is that the impact of taking an action is stable; for instance, executing the "move ahead" action will always move the agent forward by a fixed distance, perhaps with some small amount of actuator-induced noise. This assumption is limiting; an agent may encounter settings that dramatically alter the impact of actions: a move ahead action on a wet f… ▽ More

    Submitted 24 April, 2023; originally announced April 2023.

    Comments: 21 pages, 17 figures, ICLR 2023

  44. arXiv:2303.17013  [pdf

    eess.SP

    Assessing the Socio-economic Impacts of Secure Texting and Anti-Jamming Technologies in Non-Cooperative Networks

    Authors: Osoro B Ogutu, Edward J Oughton, Kai Zeng, Brian L. Mark

    Abstract: Operating securely over 5G (and legacy) infrastructure is a challenge. In non-cooperative networks, malicious actors may try to decipher, block encrypted messages, or specifically jam wireless radio systems. Such activities can disrupt operations, from causing minor inconvenience, through to fully paralyzing the functionality of critical infrastructure. While technological mitigation measures do e… ▽ More

    Submitted 10 April, 2023; v1 submitted 29 March, 2023; originally announced March 2023.

  45. arXiv:2303.16081  [pdf

    physics.class-ph physics.optics

    Overcoming losses in superlenses with synthetic waves of complex frequency

    Authors: Fuxin Guan, Kebo Zeng, Zhaoyu Nie, Xiangdong Guo, Shaojie Ma, Qing Dai, John B. Pendry, Xiang Zhang, Shuang Zhang

    Abstract: Superlenses made of plasmonic materials and metamaterials have been exploited to image features of sub-diffractional scale. However, their intrinsic losses impose a serious restriction on the imaging resolution, which is a long-standing problem that has hindered wide-spread applications of superlenses. Optical waves of complex frequency exhibiting a temporally attenuating behavior have been propos… ▽ More

    Submitted 22 March, 2023; originally announced March 2023.

    Comments: 17 pages, 3 figures

  46. arXiv:2302.06693  [pdf, ps, other

    hep-th math-ph

    Twisted Holography and Celestial Holography from Boundary Chiral Algebra

    Authors: Keyou Zeng

    Abstract: We study the Kaluza-Klein reduction of various $6d$ holomorphic theories. The KK reduction is analyzed in the BV formalism, resulting in theories that come from the holomorphic topological twist of $3d$ $\mathcal{N} = 2$ supersymmetric field theories. Effective interactions of the KK theories at the classical level can be obtained at all orders using homotopy transfer theorem. We also analyze a de… ▽ More

    Submitted 13 February, 2023; originally announced February 2023.

    Comments: 99 pages

  47. arXiv:2301.12098  [pdf, other

    physics.flu-dyn cs.LG

    Turbulence control in plane Couette flow using low-dimensional neural ODE-based models and deep reinforcement learning

    Authors: Alec J. Linot, Kevin Zeng, Michael D. Graham

    Abstract: The high dimensionality and complex dynamics of turbulent flows remain an obstacle to the discovery and implementation of control strategies. Deep reinforcement learning (RL) is a promising avenue for overcoming these obstacles, but requires a training phase in which the RL agent iteratively interacts with the flow environment to learn a control policy, which can be prohibitively expensive when th… ▽ More

    Submitted 28 January, 2023; originally announced January 2023.

  48. arXiv:2301.11586  [pdf, other

    cs.CR

    Khaos: The Impact of Inter-procedural Code Obfuscation on Binary Diffing Techniques

    Authors: Peihua Zhang, Chenggang Wu, Mingfan Peng, Kai Zeng, Ding Yu, Yuanming Lai, Yan Kang, Wei Wang, Zhe Wang

    Abstract: Software obfuscation techniques can prevent binary diffing techniques from locating vulnerable code by obfuscating the third-party code, to achieve the purpose of protecting embedded device software. With the rapid development of binary diffing techniques, they can achieve more and more accurate function matching and identification by extracting the features within the function. This makes existin… ▽ More

    Submitted 27 January, 2023; originally announced January 2023.

  49. arXiv:2212.11252  [pdf, ps, other

    math.QA hep-th math-ph math.AG

    Quadratic Duality for Chiral Algebras

    Authors: Zheng** Gui, Si Li, Keyou Zeng

    Abstract: We introduce a notion of quadratic duality for chiral algebras. This can be viewed as a chiral version of the usual quadratic duality for quadratic associative algebras. We study the relationship between this duality notion and the Maurer-Cartan equations for chiral algebras, which turns out to be parallel to the associative algebra case. We also present some explicit examples.

    Submitted 21 December, 2022; originally announced December 2022.

    Comments: 23 pages. Comments are welcome

  50. arXiv:2211.16477  [pdf

    cond-mat.str-el cond-mat.mtrl-sci cond-mat.supr-con

    Electronic nematicity without charge density waves in titanium-based kagome metal

    Authors: Hong Li, Siyu Cheng, Brenden R. Ortiz, Hengxin Tan, Dominik Werhahn, Keyu Zeng, Dirk Johrendt, Binghai Yan, Ziqiang Wang, Stephen D. Wilson, Ilija Zeljkovic

    Abstract: Layered crystalline materials that consist of transition metal atoms on a kagome network have emerged as a versatile platform to study unusual electronic phenomena. For example, in the vanadium-based kagome superconductors AV3Sb5 (where A can stand for K, Cs, or Rb) there is a parent charge density wave phase that appears to simultaneously break both the translational and the rotational symmetry o… ▽ More

    Submitted 27 July, 2023; v1 submitted 29 November, 2022; originally announced November 2022.

    Comments: This is the submitted version. Final manuscript will appear in Nature Physics

    Journal ref: Nature Physics 19, 1591 (2023)