Using predefined vector systems to speed up neural network multimillion class classification

πŸ“… 2026-04-01
πŸ“ˆ Citations: 0
✨ Influential: 0
πŸ“„ PDF
πŸ€– AI Summary
This work addresses the computational bottleneck in large-scale neural network classification, where label prediction complexity scales linearly (O(n)) with the number of classesβ€”often reaching millions. To overcome this limitation, the authors propose a geometric modeling approach in latent space based on a predefined vector system, which reformulates classification as an O(1) nearest cluster-center search by simply identifying extremal indices in the embedding vector. This method significantly reduces inference complexity without compromising training accuracy and inherently supports recognition of novel classes. Experimental results across multiple large-scale datasets demonstrate up to 11.6Γ— overall inference speedup, substantially enhancing the efficiency of ultra-large-scale classification tasks.

Technology Category

Machine Learning: Multi-class/Multi-label Learning & Extreme ClassificationComputer Vision: Large Vision ModelsNatural Language Processing: Sentence-level Semantics, Textual Inference, etc.

Application Category

Graph Algorithms and Modeling for the Web: Graph neural networks and deep learning approaches for Web-related graphsWeb Mining and Content Analysis: Large pretrained models with web dataSearch and Retrieval-Augmented AI: Efficiency and scalability of Web search engines
πŸ“ Abstract
Label prediction in neural networks (NNs) has O(n) complexity proportional to the number of classes. This holds true for classification using fully connected layers and cosine similarity with some set of class prototypes. In this paper we show that if NN latent space (LS) geometry is known and possesses specific properties, label prediction complexity can be significantly reduced. This is achieved by associating label prediction with the O(1) complexity closest cluster center search in a vector system used as target for latent space configuration (LSC). The proposed method only requires finding indexes of several largest and lowest values in the embedding vector making it extremely computationally efficient. We show that the proposed method does not change NN training accuracy computational results. We also measure the time required by different computational stages of NN inference and label prediction on multiple datasets. The experiments show that the proposed method allows to achieve up to 11.6 times overall acceleration over conventional methods. Furthermore, the proposed method has unique properties which allow to predict the existence of new classes.
Problem

Research questions and friction points this paper is trying to address.

multimillion class classification
label prediction
computational complexity
neural networks
inference acceleration
Innovation

Methods, ideas, or system contributions that make the work stand out.

latent space geometry
O(1) label prediction
predefined vector systems
multimillion-class classification
cluster center search
πŸ”Ž Similar Papers
No similar papers found.
N
Nikita Gabdullin
Joint Stock "Research and production company "Kryptonite"
I
Ilya Androsov
Joint Stock "Research and production company "Kryptonite"