Bright Vision Technologies

Bright Vision Technologies

ML Platform Engineer

Houston, Texas · Senior · Full-time

H1B sponsorship available$100k-$150kDetected 70 days ago
PythonGoRustDistributed SystemsKubernetesPlatform EngineeringMachine LearningLLMsCybersecurityIncident ResponseSystems EngineeringResearchCommunication

Stay score

odds of building a lasting career here

59Sponsors, lottery-bound
Cap-exempt (no lottery)0
Sponsors this role100
Entry-level history0
PERM / green-card track0
Lottery odds (Level III)83
Fits your clock70

Sponsors, but it's cap-subject — you still face the weighted lottery (~45% per draw at Level III). Good if you win; have a cap-exempt backup on your list.

Lottery odds assume a STEM candidate.

Personalize to your clock →

H-1B wage level

the lottery is wage-weighted — each level is one more entry

Level I · 1×
Level I$87,7761 entry
Level II$110,3442 entries
Level III$132,8913 entries
Level IV$155,4594 entries

$10,344 more$110,344 — moves this role to Level II and 2 lottery entries. That figure is inside the range the employer already advertised.

Based on the DOL prevailing wage for this occupation and worksite, a base salary of $110,344 would place this position at wage Level II. That figure is within the posted range, and I'd like to target it. This role classifies under "Software Developers" for prevailing-wage purposes.

DOL prevailing wage, 2026-27 wage year · Software Developers (15-1252) · Houston-Pasadena-The Woodlands, TX. Wage level is derived by USCIS from the offered wage, occupation and worksite; the occupation shown is inferred from the job title.

Employer immigration record

from this employer's Department of Labor filings

Files H-1B transfers

13 transfer filings in the last year, covering 13 workers. Median labor-condition decision: 7 days. An employer that already files transfers is one that can take over an existing H-1B.

Sourced from Department of Labor LCA, PERM and prevailing-wage disclosure data. Employer matching is by name, so figures may be split across an employer's legal entities. Absence of a filing means none appears in our copy of the data, not that none exists.

Community outcomes

No reports yet — be the first to help the next applicant.

About the role

  • The role focuses on the systems engineering side of AI deployment, including request routing, batching, caching, autoscaling, GPU utilization, and end-to-end observability across diverse model workloads.
  • The ideal candidate brings strong distributed systems and performance engineering expertise, has shipped serving systems at scale, and understands the trade-offs between latency, throughput, cost, and quality in ML serving.

Responsibilities

  • Optimize inference performance using continuous batching, paged attention, speculative decoding, and request multiplexing.
  • Implement multi-tenant routing, rate limiting, and quality-of-service policies across model endpoints.
  • Build autoscaling and capacity management systems that balance latency, throughput, and cost.
  • Implement caching, prompt deduplication, and response reuse strategies where appropriate.
  • Drive end-to-end observability including latency histograms, queue dynamics, GPU utilization, and error tracking.
  • Develop deployment workflows including canary releases, shadow testing, and automated rollback.
  • Operate incident response for high-availability AI services and drive durable reliability improvements.
  • Collaborate with ML and product teams to support new model releases and capability rollouts.
  • Implement security controls including request signing, content filtering, and abuse detection at the serving layer.
  • Document operational procedures, performance characteristics, and tuning guidance for internal teams.

Requirements

  • Bachelor's or Master's degree in Computer Science or a related field.
  • Six or more years of experience in distributed systems, infrastructure, or ML platform engineering.
  • Strong proficiency in Python and a systems language such as Go, Rust, or C++.
  • Deep experience operating high-throughput, low-latency services in production.
  • Hands-on experience with LLM or large model inference frameworks such as vLLM or TensorRT-LLM.
  • Strong understanding of GPU architecture, memory hierarchies, and accelerator utilization.
  • Familiarity with Kubernetes, autoscaling, and modern cloud platforms.
  • Experience with observability stacks including metrics, tracing, and structured logging.

Nice to have

  • Open-source contributions to model serving infrastructure.
  • Experience with multi-region or globally distributed AI serving.
  • Familiarity with model quantization, distillation, and compression techniques.
  • Experience supporting external-facing AI APIs at scale.

Skills

  • This is a fantastic opportunity to join an established and well-respected organization offering tremendous career growth potential.
  • ML Platform Engineer
  • 100% Remote (Continental United States)
  • In-house Bright Vision Technologies SOW engagement (no third-party client or vendor)
  • 100 K - 150 K
  • Sponsorship:
  • No new H1B sponsorship available.
  • H1B transfers welcomed for qualified candidates.
  • Full-time, direct W2 with Bright Vision Technologies (no C2C, no 1099, no third-party)
  • Long-term, multi-year, aligned to the Bright Vision SOW delivery roadmap

Compensation

  • Competitive base salary commensurate with experience, plus benefits.

Benefits

  • Design and operate model serving platforms supporting diverse workloads including LLMs, vision models, and recommendation systems.
  • This role is part of Bright Vision Technologies' in-house Statement of Work (SOW) engagement.
  • The client, end customer, and employer for this position is Bright Vision Technologies - there is no third-party client, vendor, or implementation partner involved.
  • Candidates must be willing to work directly as a full-time W2 employee of Bright Vision Technologies and contribute to our in-house SOW deliverables.
  • We are seeking a ML Platform Engineer to design, build, and operate high-performance, highly reliable inference platforms for serving large machine learning models in production.
  • Learn more about Bright Vision Technologies at www.bvteck.com.
  • We also make reasonable accommodations for applicants' and employees' religious practices and beliefs, as well as mental health or physical disability needs.
  • Bright Vision Technologies is an
  • Equal Opportunity Employer, including Disability/Veterans.

Company info

  • As we continue to grow, we're looking for a skilled ML Platform Engineer to join our dynamic team and contribute to our mission of transforming business processes through technology.
  • We are an equal opportunity employer and place a high value on diversity and inclusion at our company.

Equal opportunity

  • equal opportunity employer and place a high value on diversity and inclusion at our company.
  • Equal Employment Opportunity (EEO) Statement

Visa & Work Authorization

  • No new H1B sponsorship available.
  • H1B transfers welcomed for qualified candidates.
  • Employment Terms & Visa Policy
  • No new H1B sponsorship is available for this role.
  • However, candidates who are currently on a valid H1B visa and require a transfer are welcome to apply.
  • We will support H1B transfers for qualified candidates.

This listing is sourced directly from Bright Vision Technologies's careers page and normalized into a canonical job model.