Search | arXiv e-print repository

GeNet: A Multimodal LLM-Based Co-Pilot for Network Topology and Configuration

Authors: Beni Ifland, Elad Duani, Rubin Krief, Miro Ohana, Aviram Zilberman, Andres Murillo, Ofir Manor, Ortal Lavi, Hikichi Kenji, Asaf Shabtai, Yuval Elovici, Rami Puzis

Abstract: Communication network engineering in enterprise environments is traditionally a complex, time-consuming, and error-prone manual process. Most research on network engineering automation has concentrated on configuration synthesis, often overlooking changes in the physical network topology. This paper introduces GeNet, a multimodal co-pilot for enterprise network engineers. GeNet is a novel framewor… ▽ More Communication network engineering in enterprise environments is traditionally a complex, time-consuming, and error-prone manual process. Most research on network engineering automation has concentrated on configuration synthesis, often overlooking changes in the physical network topology. This paper introduces GeNet, a multimodal co-pilot for enterprise network engineers. GeNet is a novel framework that leverages a large language model (LLM) to streamline network design workflows. It uses visual and textual modalities to interpret and update network topologies and device configurations based on user intents. GeNet was evaluated on enterprise network scenarios adapted from Cisco certification exercises. Our results demonstrate GeNet's ability to interpret network topology images accurately, potentially reducing network engineers' efforts and accelerating network design processes in enterprise environments. Furthermore, we show the importance of precise topology understanding when handling intents that require modifications to the network's topology. △ Less

Submitted 11 July, 2024; originally announced July 2024.

arXiv:2402.01678 [pdf]

Electrical Conductance of Nanofluidic Systems subjected to Asymmetric Concentrations

Authors: Oren Lavi, Yoav Green

Abstract: A nanochannel subjected to both a potential and concentration gradient has an asymmetric current-voltage, I-V, response with three primary characteristics: the Ohmic conductance, G, the current at zero voltage, I(v=0), and the voltage at zero current, V(I=0). To date, there is no known self-consistent theory for these characteristics subject to an arbitrary concentration gradient. Here, we present… ▽ More A nanochannel subjected to both a potential and concentration gradient has an asymmetric current-voltage, I-V, response with three primary characteristics: the Ohmic conductance, G, the current at zero voltage, I(v=0), and the voltage at zero current, V(I=0). To date, there is no known self-consistent theory for these characteristics subject to an arbitrary concentration gradient. Here, we present simple expressions for each of these characteristics that have been derived self-consistently. Our findings provide insights into the underlying physics of nanofluidics systems used for water desalination and energy harvesting. △ Less

Submitted 20 January, 2024; originally announced February 2024.

Comments: 5 pages, 3 figures

arXiv:2303.01723 [pdf, other]

AI-Empowered Hybrid MIMO Beamforming

Authors: Nir Shlezinger, Mengyuan Ma, Ortal Lavi, Nhan Thanh Nguyen, Yonina C. Eldar, Markku Juntti

Abstract: Hybrid multiple-input multiple-output (MIMO) is an attractive technology for realizing extreme massive MIMO systems envisioned for future wireless communications in a scalable and power-efficient manner. However, the fact that hybrid MIMO systems implement part of their beamforming in analog and part in digital makes the optimization of their beampattern notably more challenging compared with conv… ▽ More Hybrid multiple-input multiple-output (MIMO) is an attractive technology for realizing extreme massive MIMO systems envisioned for future wireless communications in a scalable and power-efficient manner. However, the fact that hybrid MIMO systems implement part of their beamforming in analog and part in digital makes the optimization of their beampattern notably more challenging compared with conventional fully digital MIMO. Consequently, recent years have witnessed a growing interest in using data-aided artificial intelligence (AI) tools for hybrid beamforming design. This article reviews candidate strategies to leverage data to improve real-time hybrid beamforming design. We discuss the architectural constraints and characterize the core challenges associated with hybrid beamforming optimization. We then present how these challenges are treated via conventional optimization, and identify different AI-aided design approaches. These can be roughly divided into purely data-driven deep learning models and different forms of deep unfolding techniques for combining AI with classical optimization.We provide a systematic comparative study between existing approaches including both numerical evaluations and qualitative measures. We conclude by presenting future research opportunities associated with the incorporation of AI in hybrid MIMO systems. △ Less

Submitted 3 March, 2023; originally announced March 2023.

Comments: This work has been submitted to the IEEE for possible publication. Copyright may be transferred without notice, after which this version may no longer be accessible

arXiv:2301.00369 [pdf, other]

Learn to Rapidly and Robustly Optimize Hybrid Precoding

Authors: Ortal Lavi, Nir Shlezinger

Abstract: Hybrid precoding plays a key role in realizing massive multiple-input multiple-output (MIMO) transmitters with controllable cost. MIMO precoders are required to frequently adapt based on the variations in the channel conditions. In hybrid MIMO, here precoding is comprised of digital and analog beamforming, such an adaptation involves lengthy optimization and depends on accurate channel state infor… ▽ More Hybrid precoding plays a key role in realizing massive multiple-input multiple-output (MIMO) transmitters with controllable cost. MIMO precoders are required to frequently adapt based on the variations in the channel conditions. In hybrid MIMO, here precoding is comprised of digital and analog beamforming, such an adaptation involves lengthy optimization and depends on accurate channel state information (CSI). This affects the spectral efficiency when the channel varies rapidly and when operating with noisy CSI. In this work we employ deep learning techniques to learn how to rapidly and robustly optimize hybrid precoders, while being fully interpretable. We leverage data to learn iteration-dependent hyperparameter settings of projected gradient sum-rate optimization with a predefined number of iterations. The algorithm maps channel realizations into hybrid precoding settings while preserving the interpretable flow of the optimizer and improving its convergence speed. To cope with noisy CSI, we learn to optimize the minimal achievable sum-rate among all tolerable errors, proposing a robust hybrid precoding based on the projected conceptual mirror prox minimax optimizer. Numerical results demonstrate that our approach allows using over ten times less iterations compared to that required by conventional optimization with shared hyperparameters, while achieving similar and even improved sum-rate performance. △ Less

Submitted 1 January, 2023; originally announced January 2023.

Comments: 30 Pages, 9 figures

arXiv:2110.05780 [pdf, other]

We've had this conversation before: A Novel Approach to Measuring Dialog Similarity

Authors: Ofer Lavi, Ella Rabinovich, Segev Shlomov, David Boaz, Inbal Ronen, Ateret Anaby-Tavor

Abstract: Dialog is a core building block of human natural language interactions. It contains multi-party utterances used to convey information from one party to another in a dynamic and evolving manner. The ability to compare dialogs is beneficial in many real world use cases, such as conversation analytics for contact center calls and virtual agent design. We propose a novel adaptation of the edit dista… ▽ More Dialog is a core building block of human natural language interactions. It contains multi-party utterances used to convey information from one party to another in a dynamic and evolving manner. The ability to compare dialogs is beneficial in many real world use cases, such as conversation analytics for contact center calls and virtual agent design. We propose a novel adaptation of the edit distance metric to the scenario of dialog similarity. Our approach takes into account various conversation aspects such as utterance semantics, conversation flow, and the participants. We evaluate this new approach and compare it to existing document similarity measures on two publicly available datasets. The results demonstrate that our method outperforms the other approaches in capturing dialog flow, and is better aligned with the human perception of conversation similarity. △ Less

Submitted 12 October, 2021; originally announced October 2021.

Comments: EMNLP 2021, 9 pages

arXiv:2007.15547 [pdf, ps, other]

Characters of the group $\mathrm{EL}_d (R)$ for a commutative Noetherian ring $R$

Authors: Omer Lavi, Arie Levit

Abstract: Let $R$ be a commutative Noetherian ring with unit. We classify the characters of the group $\mathrm{EL}_d (R)$ provided that $d$ is greater than the stable range of the ring $R$. It follows that every character of $\mathrm{EL}_d (R)$ is induced from a finite dimensional representation. Towards our main result we classify $\mathrm{EL}_d (R)$-invariant probability measures on the Pontryagin dual gr… ▽ More Let $R$ be a commutative Noetherian ring with unit. We classify the characters of the group $\mathrm{EL}_d (R)$ provided that $d$ is greater than the stable range of the ring $R$. It follows that every character of $\mathrm{EL}_d (R)$ is induced from a finite dimensional representation. Towards our main result we classify $\mathrm{EL}_d (R)$-invariant probability measures on the Pontryagin dual group of $R^d$. △ Less

Submitted 30 July, 2020; originally announced July 2020.

Comments: 51 pages

MSC Class: 20C15; 20K30; 20H05; 20H25; 13E05; 22B05; 28D05 (Primary) 19M05; 57S30 (Secondary)

arXiv:2001.03885 [pdf, other]

Optimal Finite Homogeneous sphere approximation

Authors: Omer Lavi

Abstract: The two dimensional sphere can't be approximated by finite homogeneous spaces. We describe the optimal approximation and its distance from the sphere. The two dimensional sphere can't be approximated by finite homogeneous spaces. We describe the optimal approximation and its distance from the sphere. △ Less

Submitted 12 January, 2020; originally announced January 2020.

Comments: 12 pages 2 figures

MSC Class: 22-02

arXiv:1901.03995 [pdf, other]

Neural network gradient-based learning of black-box function interfaces

Authors: Alon Jacovi, Guy Hadash, Einat Kermany, Boaz Carmeli, Ofer Lavi, George Kour, Jonathan Berant

Abstract: Deep neural networks work well at approximating complicated functions when provided with data and trained by gradient descent methods. At the same time, there is a vast amount of existing functions that programmatically solve different tasks in a precise manner eliminating the need for training. In many cases, it is possible to decompose a task to a series of functions, of which for some we may pr… ▽ More Deep neural networks work well at approximating complicated functions when provided with data and trained by gradient descent methods. At the same time, there is a vast amount of existing functions that programmatically solve different tasks in a precise manner eliminating the need for training. In many cases, it is possible to decompose a task to a series of functions, of which for some we may prefer to use a neural network to learn the functionality, while for others the preferred method would be to use existing black-box functions. We propose a method for end-to-end training of a base neural network that integrates calls to existing black-box functions. We do so by approximating the black-box functionality with a differentiable neural network in a way that drives the base network to comply with the black-box function interface during the end-to-end optimization process. At inference time, we replace the differentiable estimator with its external black-box non-differentiable counterpart such that the base network output matches the input arguments of the black-box function. Using this "Estimate and Replace" paradigm, we train a neural network, end to end, to compute the input to black-box functionality while eliminating the need for intermediate labels. We show that by leveraging the existing precise black-box function during inference, the integrated model generalizes better than a fully differentiable model, and learns more efficiently compared to RL-based methods. △ Less

Submitted 13 January, 2019; originally announced January 2019.

Comments: Published as a conference paper at ICLR 2019

arXiv:1804.09028 [pdf, other]

Estimate and Replace: A Novel Approach to Integrating Deep Neural Networks with Existing Applications

Authors: Guy Hadash, Einat Kermany, Boaz Carmeli, Ofer Lavi, George Kour, Alon Jacovi

Abstract: Existing applications include a huge amount of knowledge that is out of reach for deep neural networks. This paper presents a novel approach for integrating calls to existing applications into deep learning architectures. Using this approach, we estimate each application's functionality with an estimator, which is implemented as a deep neural network (DNN). The estimator is then embedded into a ba… ▽ More Existing applications include a huge amount of knowledge that is out of reach for deep neural networks. This paper presents a novel approach for integrating calls to existing applications into deep learning architectures. Using this approach, we estimate each application's functionality with an estimator, which is implemented as a deep neural network (DNN). The estimator is then embedded into a base network that we direct into complying with the application's interface during an end-to-end optimization process. At inference time, we replace each estimator with its existing application counterpart and let the base network solve the task by interacting with the existing application. Using this 'Estimate and Replace' method, we were able to train a DNN end-to-end with less data and outperformed a matching DNN that did not interact with the external application. △ Less

Submitted 24 April, 2018; originally announced April 2018.

arXiv:1608.04207 [pdf, other]

Fine-grained Analysis of Sentence Embeddings Using Auxiliary Prediction Tasks

Authors: Yossi Adi, Einat Kermany, Yonatan Belinkov, Ofer Lavi, Yoav Goldberg

Abstract: There is a lot of research interest in encoding variable length sentences into fixed length vectors, in a way that preserves the sentence meanings. Two common methods include representations based on averaging word vectors, and representations based on the hidden states of recurrent neural networks such as LSTMs. The sentence vectors are used as features for subsequent machine learning tasks or fo… ▽ More There is a lot of research interest in encoding variable length sentences into fixed length vectors, in a way that preserves the sentence meanings. Two common methods include representations based on averaging word vectors, and representations based on the hidden states of recurrent neural networks such as LSTMs. The sentence vectors are used as features for subsequent machine learning tasks or for pre-training in the context of deep learning. However, not much is known about the properties that are encoded in these sentence representations and about the language information they capture. We propose a framework that facilitates better understanding of the encoded representations. We define prediction tasks around isolated aspects of sentence structure (namely sentence length, word content, and word order), and score representations by the ability to train a classifier to solve each prediction task when using the representation as input. We demonstrate the potential contribution of the approach by analyzing different sentence representation mechanisms. The analysis sheds light on the relative strengths of different sentence embedding methods with respect to these low level prediction tasks, and on the effect of the encoded vector's dimensionality on the resulting representations. △ Less

Submitted 9 February, 2017; v1 submitted 15 August, 2016; originally announced August 2016.

Showing 1–10 of 10 results for author: Lavi, O