Efficient Telecom Specific LLM: TSLAM-Mini with QLoRA and Digital Twin Data

General-purpose large language models (LLMs), despite their broad capabilities accrued from open-world data, frequently exhibit suboptimal performance when confronted with the nuanced and specialized demands inherent in real-time telecommunications applications. This investigation addresses this critical limitation through the meticulous fine-tuning of TSLAM-Mini developed by NetoAI, a compact (3.8-billion parameter) causal language model architecturally derived from Phi-4 Mini Instruct 4B. The fine-tuning regimen leverages a bespoke dataset comprising 100,000 samples, strategically engineered to address 20 pivotal telecommunications use-cases, encompassing domains such as Network Fundamentals, IP Routing, MPLS, Network Security, Automation, OSS/BSS, RAN, Mobile Core, Satellite Communications, and Ethical AI. This dataset was curated utilizing NetoAI's DigiTwin platform, enriched with granular insights from venerated network Subject Matter Experts (SMEs) and authoritative RFC documents, thereby capturing high-fidelity representations of real-world network dynamics through simulations inspired by digital twin paradigms. Employing Quantized Low-Rank Adaptation (QLoRA), a state-of-the-art Parameter Efficient Fine-Tuning (PEFT) technique, we achieved substantial training efficiency and enabled prospective deployment on resource-constrained hardware. A novel evaluation framework, predicated on a high-capacity LLM (Qwen3-235B-A22B) functioning as an automated adjudicator, was instituted to rigorously assess instruction-following fidelity and response quality across the specified telecom use-cases. Empirical results unequivocally demonstrate TSLAM-Mini's superior aptitude in telecom-centric applications, underscoring the profound efficacy of domain-specific datasets and PEFT methodologies for advancing intelligent network management.

Paper

References (17)

03“Semantic Knowledge Tuning (SK-Tuning) for efficient fine-tuning of language models,”2024 · Proceedings of the 2024 Conference on Natural Language Processing
04“Curated datasets and hyperpa-rameter optimization for specialized LLM applications,”2024 · Proceedings of the International Conference on Machine Learning Applications
05“Synthetic data generation with large language models for privacy-constrained domains,”2023 · Proceedings of the 2023 ACM Conference on Data Privacy
06“Phi-4 Mini Instruct 4B: Technical Report,”
07“Fine-tuning GPT-3.5 for load profile analysis in the energy sector,”Energy and AI
08“Synthetic data generation for digital twins in production systems,”Journal of Manufacturing Systems
09“Frameworx: eTOM, SID, and Open APIs for telecommunications,”
10“Domain adaptation of large language models for materials science using model merging,”Materials Science and Engineering: A
11“Instruction fine-tuning of smaller language models for financial tasks,”Journal of Financial Technology
12“Advances in synthetic data generation for low-resource tasks,”Computational Linguistics

Scroll for more · 5 remaining

Similar papers

© 2026 NYSGPT2525 LLC