Hacker Careers logo
Archived on

About Engineering Square


Job Description

Own RAG pipelines and agentic workflows (end-to-end accuracy); GPU-optimized model deployment (Trident-style architecture); Production MLOps including distillation, fine-tuning, and training orchestration; Data modeling for AI persona capabilities; Infrastructure responsibilities including Docker, CI/CD, logging, monitoring, and observability. Requirements: 3-12 years shipping AI/ML in production (notebooks do not count); startup or scale-up experience where you owned what you built; strong Python; bonus for C++/Java/C# in performance-critical paths; Hands-on with RAG, MLOps tooling, and GPU deployment at scale. W2 full-time. Must be authorized to work in the US. No remote.

Remote Conditions

Onsite in Bay Area, CA. No remote work option. US work authorization required.

Location

Bay Area, CA

Salary

Not Specified

Benefits

Not Specified

Tech Tags

C##Ci/cdDockerGPU deploymentJavaMLOPSPythonRag

Senior Role

Date Listed

01 May, 2026 (3 months ago)
Loading...

Share this job

This job is archived, but you can still apply.

Hiring engineers?

Reach thousands of tech candidates from the Hacker News community.

Post a Job — $99

Similar Jobs

No tags
📍 San Francisco, CA
90% match
No tags
📍 Remote
89% match
No tags
📍 Amsterdam, Berlin, Munich, Eindhoven, Ghent (EU)
88% match
No tags
📍 Amsterdam, Berlin, Munich, Eindhoven, Ghent (EU)
88% match
No tags
📍 Menlo Park, CA
88% match
No tags
📍 Palo Alto, CA or Seoul, South Korea
88% match
No tags
📍 Remote
87% match
No tags
📍 San Francisco, CA
87% match

We use cookies

We use cookies to ensure you get the best experience on our website. For more information on how we use cookies, please see our cookie policy.

By clicking “Accept”, you agree to our use of cookies.
Learn more.