--- title: 'Senior/Principal Local LLM & Generative AI Platform Engineer at Parallel Wireless' canonical: 'https://feeny.ai/job/senior-principal-local-llm-generative-ai-platform-engineer-parallel-wireless-zwzbgpfsajxd' type: 'job' last_seen: '2026-09-16' --- # Senior/Principal Local LLM & Generative AI Platform Engineer at Parallel Wireless - **Company:** Parallel Wireless - **Location:** United States - **Employment:** full-time - **Work type:** remote - **Posted:** 2026-09-09 - **Last confirmed live:** 2026-09-16 - **Apply:** https://jobs.lever.co/parallelwireless/d7aa019c-1a8b-49dd-8a3d-d5f801198dc9 ## Job description Parallel Wireless is a U.S.-based pioneer in Open RAN innovation, transforming how mobile networks are built, optimized, and powered. Through our GreenRAN™ portfolio, we help operators deliver secure, energy-efficient, automated, and flexible connectivity across 2G, 3G, 4G, 5G, and the path toward 6G. Our software-centric, hardware-agnostic approach brings intelligence into the RAN while helping customers reduce complexity and total cost of ownership. Parallel Wireless is looking for a hands-on technical leader to build and operate a secure local large-language-model platform for the company. The platform will allow engineering and business teams to use generative AI with proprietary source code, product documentation, technical standards, test artifacts, support knowledge, and other approved internal data while keeping sensitive information within company-controlled environments. This is a senior individual-contributor role spanning applied LLM engineering, platform architecture, search and data pipelines, security, and production operations. You will turn promising prototypes into a dependable internal capability: selecting and optimizing open-weight models, building permission-aware retrieval, creating reusable APIs and tools, integrating with existing engineering workflows, and establishing objective ways to measure quality, safety, latency, capacity, and business value. The successful candidate will understand that a useful enterprise LLM is more than a model and a chat interface. It requires trustworthy source grounding, strong access controls, repeatable evaluation, careful tool permissions, observable production services, and an operating model that keeps data, indexes, prompts, models, and dependencies current. You will make pragmatic build-versus-buy decisions and choose the simplest approach—search, retrieval-augmented generation (RAG), prompting, workflow automation, or model adaptation—that meets each use case. Initial use cases may include engineering knowledge discovery, source-code understanding, troubleshooting assistance, technical-document Q&A and summarization, test and log analysis, and drafting structured engineering artifacts. The platform should be extensible to additional approved use cases as needs and model capabilities evolve. What you will do: - Own the architecture and technical roadmap for a secure, reliable, and maintainable local LLM platform deployed in Parallel Wireless-controlled infrastructure. - Partner with engineering, product, support, IT, information security, legal, and domain experts to prioritize high-value use cases and translate them into measurable product and platform requirements. - Build a modular inference and model-gateway layer with stable APIs, model routing, streaming, concurrency controls, quotas, and the ability to change models or serving backends without rewriting every application. - Evaluate open-weight language, code, embedding, reranking, and, where useful, multimodal models against PW-specific tasks; document model provenance, licenses, limitations, security posture, hardware needs, and total cost of ownership. - Optimize serving across available CPU, GPU, and accelerator resources using techniques such as continuous batching, caching, parallelism, quantization, and right-sized context limits while protecting output quality. - Design and operate RAG and enterprise-search pipelines for approved repositories, wikis, tickets, standards, design documents, test results, logs, and support content, including parsing, chunking, metadata, embeddings, hybrid retrieval, reranking, freshness, citations, and deletion. - Enforce source-system permissions throughout ingestion and retrieval so that the platform never exposes content a user is not authorized to access; integrate with company identity, SSO, role-based access control, secrets management, and audit logging. - Establish versioned evaluation datasets and automated offline and online evaluation for retrieval quality, groundedness, factual accuracy, citation quality, code correctness, task completion, latency, safety, and refusal behavior. - Create release gates and reproducible regression tests for changes to models, prompts, tools, embeddings, retrieval logic, indexes, and serving configurations; support canary releases, rollback, and clear approval paths. - Implement end-to-end observability for model and agent workflows, including traces, errors, time to first token, inter-token latency, throughput, queue time, resource utilization, saturation, availability, and user feedback. - Design safe tool-calling and agent workflows with least-privilege access, sandboxing, input and output validation, bounded execution, human approval for consequential actions, and complete traceability. - Integrate the platform into the tools employees already use—such as developer environments, source-control and CI workflows, knowledge systems, ticketing systems, and internal applications—through reusable SDKs, APIs, and reference implementations. - Build the operational foundations for production use: CI/CD, configuration and model registries, backups, disaster recovery, capacity planning, dependency and vulnerability management, incident response, and lifecycle policies for models and data. - Protect proprietary and personal information through network isolation, encryption, retention controls, redaction where appropriate, secure logging, and defenses against prompt injection, data poisoning, unsafe output handling, and model-supply-chain risks. - Determine when prompt or retrieval improvements are sufficient and when parameter-efficient fine-tuning, distillation, or other adaptation is justified by measured quality gains. - Make the platform usable beyond the core AI team through documentation, examples, training, office hours, and hands-on collaboration; use telemetry and structured feedback to improve adoption and effectiveness. - Communicate architecture decisions, quality evidence, risk, capacity, and roadmap tradeoffs clearly to technical and business stakeholders. What you bring: - BSc or MSc in Computer Science, Computer Engineering, Electrical Engineering, Data Science, or a related field, or equivalent practical experience. - Typically 7+ years of hands-on experience in production software, ML platform, search, data, or infrastructure engineering, including meaningful recent experience shipping LLM-powered systems; exceptional candidates with equivalent depth are welcome. - Strong Python engineering skills and experience designing maintainable APIs, services, libraries, and data pipelines. Experience with Go, Java, or C/C++ is an advantage. - Strong understanding of transformer-based language models and production inference, including tokenization, context management, batching, KV caching, parallelism, quantization, structured output, tool calling, and common model failure modes. - Demonstrated experience building production RAG or enterprise-search systems using embeddings, vector and/or lexical search, metadata filtering, reranking, source attribution, and systematic retrieval evaluation. - Experience defining task-specific LLM evaluations using representative datasets, strong baselines, domain-expert review, automated metrics, human feedback, error analysis, and regression thresholds. - Experience deploying and operating containerized services on Linux using Docker and Kubernetes or an equivalent orchestration environment. - Practical experience with GPU-backed model serving, performance profiling, capacity planning, monitoring, and reliability engineering. - Strong knowledge of distributed-system fundamentals, authentication and authorization, API security, secrets handling, encryption, auditability, and data lifecycle controls. - Experience with Git, automated testing, CI/CD, infrastructure as code, observability, and production incident response. - Sound technical judgment about quality, security, maintainability, hardware efficiency, and total cost—not just model benchmark scores. - Ability to lead an ambiguous, cross-functional initiative, explain complex AI behavior in plain language, and help other teams ship safely on a shared platform. Nice to have: - Experience operating LLMs in on-premises, private-cloud, restricted-network, or air-gapped environments. - Hands-on experience with current inference runtimes and serving systems such as vLLM, SGLang, TensorRT-LLM, llama.cpp, Ray Serve, KServe, Triton, or equivalent technologies. - Experience optimizing inference on NVIDIA and/or AMD GPUs using CUDA, ROCm, profiling tools, tensor parallelism, pipeline parallelism, speculative decoding, prefix/KV caching, or related techniques. - Experience with model and experiment registries, LLM tracing and evaluation platforms, vector databases, hybrid-search engines, and production data-orchestration frameworks. - Experience with parameter-efficient fine-tuning methods such as LoRA/QLoRA, dataset curation, synthetic-data generation, distillation, and post-training evaluation. - Experience building code intelligence, repository-aware assistants, developer tools, or IDE and CI integrations for large C/C++ and Python codebases. - Familiarity with Active Directory or another enterprise identity provider, fine-grained document authorization, data-loss prevention, secure software supply chains, model licensing, and AI governance. - Experience red-teaming LLM or agent systems for prompt injection, sensitive-data disclosure, poisoned retrieval content, excessive agency, and insecure output handling. - Knowledge of telecommunications, 3GPP, RAN/Open RAN, cloud-native network functions, or technical-support workflows. - Experience working across heterogeneous compute platforms and making performance, energy, and TCO tradeoffs for enterprise AI workloads. - Contributions to relevant open-source AI, search, MLOps, or infrastructure projects. ## About Parallel Wireless ## Company Overview - **One-liner**: Parallel Wireless is a privately held, U.S.-based pioneer in Open RAN telecommunications, providing cloud-native, software-defined mobile network infrastructure from 2G through 5G to operators worldwide. - **Entity Type**: Private (Privately Held; multiple funding rounds including Seed, Debt Financing, Convertible Note, and Series A) - **Headquarters**: Nashua, New Hampshire, United States (300 Innovative Way, Suite #2310, Nashua, NH 03062) - **Founded**: 2012 - **Founders**: Steve Papa (Founder & CEO) ## Core Business - **Primary industry**: Telecommunications / Open RAN (Radio Access Network) infrastructure - **Target customers**: Mobile Network Operators (MNOs) globally — including MTN, Etisalat, BT EE, Tigo, and Axiata Group — with engagements across six continents and 50+ global MNOs - **Mission**: "To deliver innovative products that unlock value and disrupt the economics of wireless networks through intelligence and openness." - **Vision**: "To reimagine the wireless network so all people can be connected whenever, wherever, and however they choose." ## Products & Services - **GreenRAN™ Portfolio**: The company's flagship product line, combining energy-efficient hardware, hardware-agnostic software, and cloud-first automation to reduce power consumption, lower total cost of ownership (TCO), and accelerate sustainable 5G deployment. Includes fully automated outdoor and indoor coverage/capacity solutions that are software-upgradable to 5G. - **ALL G O-RAN Software Platform**: A cloud-native, open, secure, and intelligent RAN architecture supporting 5G/4G/3G/2G interoperability, designed to reduce network complexity and deployment costs. Fully compliant and interoperable OpenRAN architecture, built from the ground up to be software-defined. - **Network Automation & Services**: Professional services and network engineering support (deployment, integration, validation) delivered alongside the RAN platform, including AI/ML-driven integration and validation capabilities. ## Market Standing - **Valuation/Market Cap**: Not disclosed (private company) - **Key Metric**: Annual Revenue of approximately $51M (per LinkedIn company data); Total Funding of approximately $10.6M across multiple rounds - **Notable Funding Rounds**: - Seed Round (2014): 1 investor - Debt Financing (Feb 2016): $8.8M, 1 investor - Convertible Note (Mar 2016): $1.8M, 1 investor - Series A (2018): 1 investor - **Notable Investors/Partners/Customers**: MTN, Etisalat, BT EE, Tigo, Axiata Group (operator collaborations); engaged with 50+ global MNOs - **Recognition**: 74+ industry awards - **Growth Signals**: Headcount of 630 employees, up +16.5% YoY (+117 people); workforce distributed across 17 countries including Israel, India, United States, United Kingdom, South Korea, Tanzania, Turkey, Nigeria, and Kenya; LinkedIn followers of 89,561 (+39.8% yearly growth); open roles across sales, engineering (4G/5G stack), and AI-driven integration/validation ## Competitive Advantages - **Open RAN Pioneer**: One of the earliest and most prominent companies in the Open RAN movement, with a track record across six continents and a mature, deployed footprint with major MNOs. - **Software-Defined, Hardware-Agnostic Architecture**: The ALL G O-RAN platform reduces vendor lock-in and allows operators to use any hardware, differentiating on software intelligence and automation rather than proprietary hardware. - **Energy Efficiency & Sustainability**: GreenRAN™ portfolio positions the company at the intersection of 5G expansion and energy-cost reduction — a pressing operator priority. - **Full Generational Coverage**: Unique ability to evolve networks from 2G/3G/4G to 5G via software upgrades, lowering customer TCO and simplifying migration paths. - **Global Engineering Footprint**: Distributed R&D across Israel (largest hub, 243 employees) and India (230 employees), enabling cost-effective, around-the-clock development capacity. ## Strategic Focus - **Expanding Open RAN adoption**: Driving the industry shift from viewing Open RAN as a finished standard to a foundation for continuous, software-driven innovation. - **AI/ML integration**: Hiring for AI-driven integration and validation engineers and AI/ML-expert tech leads in 4G/5G stack development, signaling a push to embed intelligence into the RAN. - **Sustainable 5G expansion**: Emphasizing energy-efficient RAN deployments to help operators meet sustainability goals while lowering TCO. - **Global commercial growth**: Active sales hiring across Africa, Southeast Asia, and the Pacific region (Fiji, New Zealand), indicating expansion into emerging and high-growth mobile markets. ## Why Work Here - **Pioneering, mission-driven work**: Employees have the opportunity to help "disrupt, challenge, and lead the future of telecommunications" within the global Open RAN movement, developing products from 2G to 5G and beyond. - **Global, distributed culture**: Workforce spans 17 countries, with major hubs in Israel, India, the US, and the UK — offering a genuinely international work environment and a global mobility program. - **Flexible/remote working**: The company explicitly supports flexible and remote working arrangements, plus paid time off to rest and recharge, and time off to give back to the community. - **Career growth and development**: Rapid career growth opportunities, tailored individual development plans delivered regionally, and a "Servant Leadership" philosophy where leadership is vested in employee success. - **Recognition and rewards**: Spot Awards for celebrating wins; competitive total rewards package with long-term wealth creation opportunities described as "unmatched in the industry." - **Engineering culture**: Heavily technical organization (~55% of employees in technical roles), with challenging assignments, collaborative teamwork, and an emphasis on innovation, openness, andamos customer success. - **⚠️ Glassdoor/LinkedIn review caution**: Employer rating is 3.5/5.0 based on 240 reviews, with Work-Life Balance 3.2, Compensation 3.3, Culture 3.2, and Career Development 3.0 — candidate should evaluate these areas carefully during interviews. - **Open roles (sample)**: Director of Solution Sales Engineering (Pacific), Principal Systems Engineer (RF Communications & Sensing), Account Manager OpenRAN (Africa), Junior Engineer RT 5G Stack, Sr. Engineer RT 5G Stack, 5G/LTE Network Engineer I, Sales Director / Customer Executive (Southeast Asia, Pacific) — see the Lever careers page for the full list. ## Sources 1. [parallelwireless.com — Who We Are](https://www.parallelwireless.com/company/who-we-are/) 2. [parallelwireless.com — Careers](https://www.parallelwireless.com/careers/) 3. [jobs.lever.co — Parallel Wireless Open Positions](https://jobs.lever.co/parallelwireless) 4. [linkedin.com — Parallel Wireless Company Profile](https://linkedin.com/company/parallel-wireless-inc) 5. [cbinsights.com — Parallel Wireless Company Profile](https://www.cbinsights.com/company/parallel-wireless) ## Other roles at Parallel Wireless - [Technical Lead, Artificial Intelligence](https://feeny.ai/job/technical-lead-artificial-intelligence-parallel-wireless-kfar-saba-yasyzvdx9znq) — Kfar Saba, Israel - [AI Engineering Team Manager](https://feeny.ai/job/ai-engineering-team-manager-parallel-wireless-kfar-saba-s0ewkyb287nz) — Kfar Saba, Israel - [Senior Data Scientist](https://feeny.ai/job/senior-data-scientist-parallel-wireless-kfar-saba-s7rbcqca06kz) — Kfar Saba, Israel - [RAN Digital Twin Engineer](https://feeny.ai/job/ran-digital-twin-engineer-parallel-wireless-kfar-saba-70tkcmakgc16) — Kfar Saba, Israel - [Senior PHY Software Engineer – 2G/4G/5G](https://feeny.ai/job/senior-phy-software-engineer-2g-4g-5g-parallel-wireless-bengaluru-2hzg4q5kyczr) — Bengaluru, India - [Senior/Principal RAN Digital Twin & AI Simulation Engineer](https://feeny.ai/job/senior-principal-ran-digital-twin-ai-simulation-engineer-parallel-wireless-rkb328gj6ksc) — United States - [PHY Algorithms Senior Engineer - AI/ML](https://feeny.ai/job/phy-algorithms-senior-engineer-ai-ml-parallel-wireless-kfar-saba-x2jhqj43y7mx) — Kfar Saba, Israel - [Engineer I, RU Validation -Automation ver (Testing)](https://feeny.ai/job/engineer-i-ru-validation-automation-ver-testing-parallel-wireless-bengaluru-t4hwyd0gywas) — Bengaluru, India - [RF Planning & Design Engineer](https://feeny.ai/job/rf-planning-design-engineer-parallel-wireless-pune-v4w472d3da3a) — Pune, India - [Senior Automation Engineer - North](https://feeny.ai/job/senior-automation-engineer-north-parallel-wireless-kinneret-xfcwy0stjbnt) — Kinneret, Israel