Relomote
Remote JobsRelocation Jobs
Add companySaved
Relomote

Relomote is a job board for remote, hybrid, and relocation jobs — every listing AI-classified for the countries it actually hires from, or the visa and relocation support it offers.

LinkedInCrunchbase

Remote jobs by category

  • Remote Engineering & Development jobs
  • Remote Customer Support jobs
  • Remote Design jobs
  • Remote Marketing jobs
  • Remote Sales jobs
  • Remote Product jobs
  • Remote Data & Analytics jobs
  • Remote People & Talent jobs
  • Remote Writing & Content Creation jobs
  • Remote Finance jobs
  • Remote Legal & Compliance jobs
  • Remote Operations & Admin jobs
  • Remote Data Entry jobs
  • Remote Virtual Assistant jobs
  • Remote Education/Training jobs
  • Remote Healthcare/Clinical jobs
  • Remote Other jobs

Remote jobs by location

  • Work from anywhere jobs
  • Remote jobs in Africa
  • Remote jobs in Asia
  • Remote jobs in Europe
  • Remote jobs in Latin America
  • Remote jobs in Middle East
  • Remote jobs in North America
  • Remote jobs in Oceania
  • All remote jobs →

Relocation & visa sponsorship

  • Visa sponsorship jobs
  • Relocation package jobs
  • Relocate to Europe
  • Relocate to Germany
  • Relocate to Netherlands
  • Relocate to Spain
  • Relocate to Portugal
  • Relocate to Greece
  • Relocate to United Kingdom
  • Relocate to Canada
  • Relocate to Australia
  • Relocate to Sweden
  • Relocate to Switzerland
  • Relocate to Japan
  • Relocate to United Arab Emirates
  • All relocation jobs →

© 2026 RelomoteAboutPrivacyTerms

Contact [email protected] · Built by Mahmoud

Relomote
Remote JobsRelocation Jobs
Add companySaved
NextGen Federal Systems logo

Edge AI/Model Optimization Engineer

NextGen Federal Systems
Posted May 28, 2026, 12:27 AM UTC
🇺🇸United States🏢Hybrid📁Data & Analytics
Is this job info correct?

NextGen is seeking a highly motivated and technically skilled Edge AI/Model Optimization Engineer to support the deployment, optimization, and sustainment of AI and agentic AI capabilities within edge and tactical computing environments. This role focuses on evaluating, tuning, benchmarking, and operationalizing Large Language Models (LLMs), embedding models, and AI inference services for constrained hardware platforms, including the X9 Spider Mission Computer architecture and other edge compute systems supporting operational missions using ReadiChat. ReadiChat is a mission-focused, agentic AI platform designed to help organizations build, deploy, govern, and scale specialized AI agents for operational workflows. It combines AI agents, workflow orchestration, grounded knowledge, testing frameworks, and enterprise controls into a single collaborative workspace. The ideal candidate will possess expertise in AI model optimization, GPU-enabled edge computing, runtime performance tuning, and operational AI deployment. This role requires close collaboration with AI engineers, systems integrators, mission stakeholders, and operational users to ensure AI-enabled capabilities remain performant, reliable, and mission-effective within disconnected, degraded, intermittent, and low-bandwidth environments. Responsibilities Evaluate candidate Large Language Models (LLMs), embedding models, and AI inference solutions for quality, latency, memory utilization, reliability, and operational performance on embedded GPU-enabled edge compute platforms, including the X9 Spider Mission Computer architecture. Tune and optimize AI model runtime configurations for edge deployment, including quantization strategies, batching configurations, context window sizing, cache behavior, inference scheduling, and GPU memory utilization specific to operational edge hardware environments. Collaborate with customer stakeholders to assess mission requirements and evaluate alternative edge compute platforms when operational demands exceed X9 Spider capabilities or when cost, performance, power, size, weight, or thermal tradeoffs require additional analysis. Benchmark agentic AI workflows, inference pipelines, and model-serving architectures against target hardware constraints and operational performance thresholds. Recommend model-selection, runtime, and configuration tradeoffs balancing mission effectiveness, latency, throughput, resource utilization, reliability, and operational sustainability. Build and maintain repeatable performance and stress-testing frameworks for evaluating latency, throughput, tool-call overhead, failover behavior, degraded-resource conditions, and disconnected operational scenarios on edge compute platforms. Package, deploy, validate, and sustain local model-serving components and inference services to support reliable operation within tactical and edge environments. Collaborate with agent engineers, AI developers, and integration teams to validate that agent behavior, workflow reliability, and operational outcomes remain acceptable following model compression, quantization, runtime optimization, or hardware configuration changes. Support deployment, troubleshooting, optimization, and sustainment activities for AI-enabled applications operating in edge, airborne, tactical, or disconnected operational environments. Train customer technical personnel on supported model profiles, operational constraints, runtime tuning considerations, deployment limitations, troubleshooting procedures, and platform sustainment best practices. Maintain technical documentation, benchmarking results, model validation reports, deployment procedures, optimization baselines, configuration guides, and operational support materials. Support DevSecOps and CI/CD activities associated with AI model packaging, deployment automation, runtime validation, and operational release processes. Required Qualifications Bachelor’s degree in Computer Science, Electrical Engineering, Computer Engineering, Data Science, Artificial Intelligence, or related technical discipline. 5+ years of experience supporting AI/ML deployment, model optimization, edge computing, GPU acceleration, or AI inference operations. Experience deploying and optimizing LLMs, embedding models, or AI inference pipelines within resource-constrained or edge-compute environments. Experience with GPU-enabled systems and inference optimization technologies such as CUDA, TensorRT, ONNX Runtime, vLLM, Ollama, or equivalent platforms. Experience tuning AI runtime configurations including quantization, batching, caching, and memory optimization techniques. Experience benchmarking AI models and operational workflows against hardware performance constraints. Experience with Linux-based systems, containerized deployments, and orchestration technologies such as Docker and Kubernetes. Familiarity with Python and AI/ML deployment frameworks commonly used for edge inference and operational AI systems. Strong analytical, troubleshooting, and performance optimization skills. Ability to communicate technical findings and operational tradeoffs effectively to technical and non-technical stakeholders. Active Security Clearance is required Desired Qualifications Experience supporting tactical, airborne, or mission-command edge computing environments. Familiarity with X9 Spider Mission Computer architectures or similar embedded GPU-enabled mission systems. Experience supporting AI-enabled workflows within NGC2, AIDP, EMSCO, Lattice, or related operational ecosystems. Experience with model quantization techniques such as INT8, FP16, GGUF, GPTQ, AWQ, or similar optimization approaches. Familiarity with disconnected, degraded, intermittent, and low-bandwidth (DDIL) operational environments. Experience with hardware evaluation and performance trade studies for operational edge compute systems. About NextGen: NextGen Federal Systems is an innovative technology and professional services provider specializing in advanced software solutions and comprehensive mission and business support services. We work in close collaboration with our customers to truly understand their business and mission goals. Our approach is to design, build, implement, and manage solutions that measurably improve our client’s organizational performance. We have established and foster a corporate culture where we: Treat employees with fairness and respect regardless of their position, sexual identity, race, or tenure. Communicate the importance of our mission and our employees’ contributions to it, ensuring they understand how their job role contributes to the greater good. Openly promote and communicate our ideas for change and adaptability. Strive to achieve results as an organization. Hold employees accountable to their commitments and provide incentives that encourage positive and productive behaviors. Value the talents and contributions of our employees as the key factor for our success. Create an environment where people can engage at all levels. Encourage people to take risks and allow them to make mistakes. Equal Opportunity Employer/Protected Veterans/Individuals with Disabilities. RefID: A01

Similar jobs

Similar jobs

Nvidia logo

Senior Data Center Performance Engineer - Benchmarking and Optimization

Nvidia

🇺🇸United States14 hours ago
MA

Senior SoC Power Analysis & Optimization Engineer

MatX

🇺🇸United StatesYesterday
Nvidia logo

Senior Deep Learning Software Engineer, Inference and Model Optimization

Nvidia

🇺🇸United States2 days ago
Bright Vision Technologies logo

AI Optimization Engineer

Bright Vision Technologies

🇺🇸United States4 days ago
Bright Vision Technologies logo

Model Optimization Engineer

Bright Vision Technologies

🇺🇸United States4 days ago
Syniverse logo

Sr Infrastructure and Capacity Optimization Engineer

Syniverse

🇺🇸United States4 days ago