Edge ML Engineer

Bright Vision Technologies

  • Raleigh, North Carolina, United States
  • Remote
  • $100,000 - $150,000 a year
  • Posted Aug 15, 2026
Sign up — let your agent apply Sign in

BestApply tailors your resume and applies for you.

C++ProfilingModel CompressionOn-device PrivacyPruningPythonOn-device securityEdge inference frameworksML Model DeploymentQuantizationEmbedded Hardware ArchitectureMobile Hardware Architecture

Job description

About the role

The Edge ML Engineer will design, optimize, and deploy machine learning models for resource‑constrained edge devices such as mobile platforms, embedded systems, and specialized accelerators, working remotely to ship reliable AI capabilities outside the data center.

About the company

Bright Vision Technologies is a technology consulting and software development firm that delivers cloud, AI, data, and enterprise solutions across the United States, positioned as an established, well‑respected organization with strong career growth potential.

Requirements

  • Bachelor’s or Master’s degree in Computer Science, Computer Engineering, or related field.
  • Six or more years of experience in ML engineering with significant edge or mobile AI work.
  • Strong proficiency in Python and C++.
  • Hands‑on experience with model compression, quantization, and pruning techniques.
  • Experience with at least one major edge inference framework.
  • Solid understanding of mobile and embedded hardware architectures.
  • Experience deploying ML models to production on mobile or embedded platforms.
  • Strong performance engineering and profiling skills.
  • Familiarity with on‑device privacy and security considerations.
  • Strong communication and cross‑functional collaboration skills.
  • Experience with custom NPU or DSP toolchains (preferred).
  • Familiarity with federated learning or on‑device personalization (preferred).
  • Exposure to safety‑critical or industrial edge deployments (preferred).
  • Open‑source contributions to edge AI frameworks (preferred).
  • Experience optimizing LLMs for on‑device inference (preferred).