• Skip to primary navigation
  • Skip to main content
  • Skip to footer

Side Hustles

Side Hustles

Side Hustles For All

  • Best Side Hustles
    • Woman sitting on a pile of coins and working on a laptop surrounded by icons representing different side hustle ideas

      31 Best Side Hustles to Earn Extra Money in 2026

    • Bicycle courier delivering food for their side hustle.

      What Is a Side Hustle?

    • Remote worker sitting at his desk making money from home

      18 Ways to Make Money from Home (Online and Offline Jobs)

    • By Category
      • Arts & Crafts
      • Business Services
      • Caregiving
      • Creative Services
      • Digital Freelance Services
      • View All
    • By Lifestyle
      • I’m introverted
      • I’m a man
      • I’m a woman
      • I’m a stay-at-home mom
      • I’m unique
      • View All
    • By Profession
      • Artists & Creatives
      • Musicians
      • Nurses
      • Physicians
      • Teachers
      • View All
    • By Age Group
      • College Students
      • Teens
      • Age 50+
      • Seniors
      • View All
    • By Skills & Interests
      • Get Paid to Lose Weight
      • Get Paid to Play Games
      • Get Paid to Read
      • Get Paid to Sleep
      • Get Paid to Travel
      • View All
  • Best Gig Apps
    • Freelance worker popping out of a phone screen and considering gig apps on the App Store and Google Play

      Top 6 Gig Apps to Make Real Cash in 2026

    • Smartphone surrounded by the icons of different money-making apps

      Top 10 Best Money-Making Apps to Try in 2026

    • two teenagers using job apps on a phone and laptop

      19 Job Apps for Teens to Find Jobs and Make Money

    • By Gig Type
      • Cashback
      • Data Entry
      • Delivery
      • Games
      • Product Testing
      • View All
    • By Payment Method
      • Bingo Games that Pay to Cash App
      • Games that Pay Real Money
      • Games that Pay to Cash App
      • Games that Pay via PayPal
      • Surveys that Pay to Cash App
      • View All
    • By Benefits
      • $20 Signup Bonuses
      • $25 Signup Bonuses
      • $50 Signup Bonuses
      • Best Signup Bonuses
      • Instant Signup Bonuses
      • View All
    • By Skills & Interests
      • Driving
      • Losing Weight
      • Playing Games
      • Product Testing
      • Watching Ads
      • View All
  • Job Hunting
    • Freelance worker browsing a job post on a freelance job board.

      23 Job Boards You Can Use to Find Remote Work

    • Freelance writer sitting at her laptop working on a project

      15 Best Remote Jobs That Require No Paid Work Experience

    • Teenager sitting at laptop working an online job

      14 Online Jobs for Teens (With No Experience)

    • Freelancing
      • Freelance Writing Sites
      • Freelance Writing Job Boards
      • Freelance Writing Platforms
      • View More
    • Gig & Shift Work
      • Gig Work Apps
      • On-Demand Work Apps
      • Shift Work Apps
      • View More
    • GPT (Get Paid To)
      • Microtasking
      • Product Testing
      • Survey Taking
      • View More
    • Remote Working
      • Best Remote Job Boards
      • Top 15 Remote Jobs
      • View More
  • Job Board
    • Work Schedule
      • Part-Time Jobs
      • Per-Diem Jobs
      • Go Search
    • Work Environment
      • Hybrid Jobs
      • Remote Jobs
      • Go Search
    • Employment Type
      • Contractor Jobs
      • Internship Jobs
      • Temporary Jobs
      • Go Search
    • Job Title
      • Accounting Jobs
      • Data Entry Jobs
      • Nursing Jobs
      • Online Teaching Jobs
      • Software Engineer Jobs
      • Go Search
    • State
      • California Jobs
      • Florida Jobs
      • New York Jobs
      • Pennsylvania Jobs
      • Texas Jobs
      • Go Search
    • City
      • Chicago, IL
      • Houston, TX
      • Los Angeles, CA
      • New York City, NY
      • Phoenix, AZ
      • Go Search

Home Flexible Job Board Senior/Principal Local LLM & Generative AI Platform Engineer

Salary Unstated 16d ago

Senior/Principal Local LLM & Generative AI Platform Engineer

Boost your chances before you apply.

  • ✨ Apply 10x Faster Free

    It takes 30+ tailored applications to land jobs like this one. We'll help you get that done in 1 hour.

    No Credit Card Required

  • Proceed to Application Go directly to the company's job page to apply.
Logo

Parallel Wireless

US

Full-time Permanent Remote

✨ Apply 10x Faster

Analyze your resume for missing keywords, then one-click optimize it. Don't be anything less than a 100% match candidate.

Free

No Credit Card Required

Summary

You will architect and operate a secure, scalable local LLM platform to support engineering and business workflows. This involves building inference gateways, RAG pipelines, and observability tools while ensuring data security and model performance.

Job Description

Parallel Wireless is a U.S.-based pioneer in Open RAN innovation, transforming how mobile networks are built, optimized, and powered. Through our GreenRAN™ portfolio, we help operators deliver secure, energy-efficient, automated, and flexible connectivity across 2G, 3G, 4G, 5G, and the path toward 6G.
Our software-centric, hardware-agnostic approach brings intelligence into the RAN while helping customers reduce complexity and total cost of ownership. 

Parallel Wireless is looking for a hands-on technical leader to build and operate a secure local large-language-model platform for the company. The platform will allow engineering and business teams to use generative AI with proprietary source code, product documentation, technical standards, test artifacts, support knowledge, and other approved internal data while keeping sensitive information within company-controlled environments. 

This is a senior individual-contributor role spanning applied LLM engineering, platform architecture, search and data pipelines, security, and production operations. You will turn promising prototypes into a dependable internal capability: selecting and optimizing open-weight models, building permission-aware retrieval, creating reusable APIs and tools, integrating with existing engineering workflows, and establishing objective ways to measure quality, safety, latency, capacity, and business value. 

The successful candidate will understand that a useful enterprise LLM is more than a model and a chat interface. It requires trustworthy source grounding, strong access controls, repeatable evaluation, careful tool permissions, observable production services, and an operating model that keeps data, indexes, prompts, models, and dependencies current. You will make pragmatic build-versus-buy decisions and choose the simplest approach—search, retrieval-augmented generation (RAG), prompting, workflow automation, or model adaptation—that meets each use case. 

Initial use cases may include engineering knowledge discovery, source-code understanding, troubleshooting assistance, technical-document Q&A and summarization, test and log analysis, and drafting structured engineering artifacts. The platform should be extensible to additional approved use cases as needs and model capabilities evolve. 

\n

What you will do:

  • Own the architecture and technical roadmap for a secure, reliable, and maintainable local LLM platform deployed in Parallel Wireless-controlled infrastructure. 

  • Partner with engineering, product, support, IT, information security, legal, and domain experts to prioritize high-value use cases and translate them into measurable product and platform requirements. 

  • Build a modular inference and model-gateway layer with stable APIs, model routing, streaming, concurrency controls, quotas, and the ability to change models or serving backends without rewriting every application. 

  • Evaluate open-weight language, code, embedding, reranking, and, where useful, multimodal models against PW-specific tasks; document model provenance, licenses, limitations, security posture, hardware needs, and total cost of ownership. 

  • Optimize serving across available CPU, GPU, and accelerator resources using techniques such as continuous batching, caching, parallelism, quantization, and right-sized context limits while protecting output quality. 

  • Design and operate RAG and enterprise-search pipelines for approved repositories, wikis, tickets, standards, design documents, test results, logs, and support content, including parsing, chunking, metadata, embeddings, hybrid retrieval, reranking, freshness, citations, and deletion. 

  • Enforce source-system permissions throughout ingestion and retrieval so that the platform never exposes content a user is not authorized to access; integrate with company identity, SSO, role-based access control, secrets management, and audit logging. 

  • Establish versioned evaluation datasets and automated offline and online evaluation for retrieval quality, groundedness, factual accuracy, citation quality, code correctness, task completion, latency, safety, and refusal behavior. 

  • Create release gates and reproducible regression tests for changes to models, prompts, tools, embeddings, retrieval logic, indexes, and serving configurations; support canary releases, rollback, and clear approval paths. 

  • Implement end-to-end observability for model and agent workflows, including traces, errors, time to first token, inter-token latency, throughput, queue time, resource utilization, saturation, availability, and user feedback. 

  • Design safe tool-calling and agent workflows with least-privilege access, sandboxing, input and output validation, bounded execution, human approval for consequential actions, and complete traceability. 

  • Integrate the platform into the tools employees already use—such as developer environments, source-control and CI workflows, knowledge systems, ticketing systems, and internal applications—through reusable SDKs, APIs, and reference implementations. 

  • Build the operational foundations for production use: CI/CD, configuration and model registries, backups, disaster recovery, capacity planning, dependency and vulnerability management, incident response, and lifecycle policies for models and data. 

  • Protect proprietary and personal information through network isolation, encryption, retention controls, redaction where appropriate, secure logging, and defenses against prompt injection, data poisoning, unsafe output handling, and model-supply-chain risks. 

  • Determine when prompt or retrieval improvements are sufficient and when parameter-efficient fine-tuning, distillation, or other adaptation is justified by measured quality gains. 

  • Make the platform usable beyond the core AI team through documentation, examples, training, office hours, and hands-on collaboration; use telemetry and structured feedback to improve adoption and effectiveness. 

  • Communicate architecture decisions, quality evidence, risk, capacity, and roadmap tradeoffs clearly to technical and business stakeholders. 

What you bring:

  • BSc or MSc in Computer Science, Computer Engineering, Electrical Engineering, Data Science, or a related field, or equivalent practical experience. 

  • Typically 7+ years of hands-on experience in production software, ML platform, search, data, or infrastructure engineering, including meaningful recent experience shipping LLM-powered systems; exceptional candidates with equivalent depth are welcome. 

  • Strong Python engineering skills and experience designing maintainable APIs, services, libraries, and data pipelines. Experience with Go, Java, or C/C++ is an advantage. 

  • Strong understanding of transformer-based language models and production inference, including tokenization, context management, batching, KV caching, parallelism, quantization, structured output, tool calling, and common model failure modes. 

  • Demonstrated experience building production RAG or enterprise-search systems using embeddings, vector and/or lexical search, metadata filtering, reranking, source attribution, and systematic retrieval evaluation. 

  • Experience defining task-specific LLM evaluations using representative datasets, strong baselines, domain-expert review, automated metrics, human feedback, error analysis, and regression thresholds. 

  • Experience deploying and operating containerized services on Linux using Docker and Kubernetes or an equivalent orchestration environment. 

  • Practical experience with GPU-backed model serving, performance profiling, capacity planning, monitoring, and reliability engineering. 

  • Strong knowledge of distributed-system fundamentals, authentication and authorization, API security, secrets handling, encryption, auditability, and data lifecycle controls. 

  • Experience with Git, automated testing, CI/CD, infrastructure as code, observability, and production incident response. 

  • Sound technical judgment about quality, security, maintainability, hardware efficiency, and total cost—not just model benchmark scores. 

  • Ability to lead an ambiguous, cross-functional initiative, explain complex AI behavior in plain language, and help other teams ship safely on a shared platform. 

Nice to have:

  • Experience operating LLMs in on-premises, private-cloud, restricted-network, or air-gapped environments. 

  • Hands-on experience with current inference runtimes and serving systems such as vLLM, SGLang, TensorRT-LLM, llama.cpp, Ray Serve, KServe, Triton, or equivalent technologies. 

  • Experience optimizing inference on NVIDIA and/or AMD GPUs using CUDA, ROCm, profiling tools, tensor parallelism, pipeline parallelism, speculative decoding, prefix/KV caching, or related techniques. 

  • Experience with model and experiment registries, LLM tracing and evaluation platforms, vector databases, hybrid-search engines, and production data-orchestration frameworks. 

  • Experience with parameter-efficient fine-tuning methods such as LoRA/QLoRA, dataset curation, synthetic-data generation, distillation, and post-training evaluation. 

  • Experience building code intelligence, repository-aware assistants, developer tools, or IDE and CI integrations for large C/C++ and Python codebases. 

  • Familiarity with Active Directory or another enterprise identity provider, fine-grained document authorization, data-loss prevention, secure software supply chains, model licensing, and AI governance. 

  • Experience red-teaming LLM or agent systems for prompt injection, sensitive-data disclosure, poisoned retrieval content, excessive agency, and insecure output handling. 

  • Knowledge of telecommunications, 3GPP, RAN/Open RAN, cloud-native network functions, or technical-support workflows. 

  • Experience working across heterogeneous compute platforms and making performance, energy, and TCO tradeoffs for enterprise AI workloads. 

  • Contributions to relevant open-source AI, search, MLOps, or infrastructure projects. 

\n

About the company

Parallel Wireless

At Parallel Wireless, we believe that software has the power to unleash amazing opportunities for the world. We disrupt the ways wireless networks are built and operated. We are reimagining how hardware, software and the cloud work together to change deployment economics for our customers. Our ALL G O-RAN software platform forms an open, secure and intelligent RAN architecture to deliver wireless connectivity, so all people can be connected whenever, wherever, and however they choose. We are engaged with over 50 global MNOs and have been recognized with over 74 industry awards. At the core of what we do is our team of Reimaginers who value innovation, collaboration, openness and customer success. For more information, visit: www.parallelwireless.com.

Founded

2012

Company size

501-1,000 employees

Industry

Telecommunications

Org type

Privately Held

Headquarters

Nashua, NH

Apply Now

About the company

Parallel Wireless

At Parallel Wireless, we believe that software has the power to unleash amazing opportunities for the world. We disrupt the ways wireless networks are built and operated. We are reimagining how hardware, software and the cloud work together to change deployment economics for our customers. Our ALL G O-RAN software platform forms an open, secure and intelligent RAN architecture to deliver wireless connectivity, so all people can be connected whenever, wherever, and however they choose. We are engaged with over 50 global MNOs and have been recognized with over 74 industry awards. At the core of what we do is our team of Reimaginers who value innovation, collaboration, openness and customer success. For more information, visit: www.parallelwireless.com.

Founded

2012

Company size

501-1,000 employees

Industry

Telecommunications

Org type

Privately Held

Headquarters

Nashua, NH

Footer

sidehustles.com
Facebook Twitter Instagram LinkedIn Reddit TikTok YouTube

Show Me The Money

  • Side Hustle Basics
  • Side Hustle Job Board (Remote & Part-Time Jobs)
  • Gig App Reviews
  • Job Hunting
  • Manage Your Money
  • The Gig Apple: News & Events

Company

  • About Us
  • Contact Us
  • Become a Contributor
  • Advertising & Sponsorships
  • Partner With Us
  • Editorial Guidelines

Side Hustles © All rights reserved

  • Privacy Policy
  • Terms of Service

Thanks for using our free job board

Your review would mean a lot to us.

If you love that we're just giving away remote jobs for free with no paywall, please spread the word. (You will need to create an account on Trustpilot, for which we'll be eternally grateful.) Good luck out there!

Leave a Review Not yet. Send me to the job post.