Micro1
Software Engineer (Human Data Platforms)
Posted
2 weeks ago
Experience
2+ Years
Salary
$140K - $220K/yr
Deadline
Closed
Job Summary
The Software Engineer (Human Data Platforms) builds and ships end-to-end features across the entire software stack to support AI model evaluation and data pipelines. Core daily responsibilities include architecting backend APIs with Python/FastAPI; crafting intuitive, high-performance UIs using React, Next.js, and TypeScript; engineering secure cloud document storage pipelines (AWS S3); parsing complex nested JSON schemas; and designing scalable OCR and search systems for dense data structures.
Foundational Capabilities & Prerequisites
- Stack Versatility: Proven, professional track record as a full-stack builder with deep, production-grade development experience across both Python (FastAPI) and React.
- Frontend Mastery: Strong engineering familiarity with modern user interface tools, specifically TypeScript, Next.js, and state-management frameworks.
- Software Fundamentals: Exceptional command of core computer science fundamentals, including asynchronous programming models, clean REST/GraphQL API design, robust relational data modeling, and optimized SQL (PostgreSQL) indexing.
- Cloud Architecture: Direct experience provisioning, scaling, and managing cloud infrastructure platforms, with specific proficiency in AWS (or comparable GCP/Azure environments).
- Pipeline Engineering: Demonstrable experience building, debugging, and scaling data processing systems, message queues, or automated asset validation pipelines.
- Operational Persona: High personal ownership, low ego, exceptional analytical problem-solving skills, and a relentless drive for fast execution.
Preferred Technical Multipliers
- AI/LLM Exposure: Hands-on experience building interfaces or data loaders that interact with Large Language Models, prompt engines, or machine learning model evaluation suites.
- Document Processing: Prior work building optical character recognition (OCR) systems, text extraction engines, elastic search indexing, or custom PDF generation/annotation tools.
- Secure Data Governance: Experience implementing rigid role-based access control (RBAC), multi-tenant data isolation, data loss prevention (DLP) frameworks, or compliance-grade encryption pipelines.
Technical Architecture Focus
In this role, your daily development sprints will tackle the intersection of big-data processing and intuitive developer tool design:
- Secure Object Storage Orchestration: Building dynamic pre-signed URL architectures and granular permission layers across AWS S3 buckets to enforce absolute data isolation with zero leakage.
- Schema Enforcement Engines: Designing high-performance validation layers capable of parsing, cleaning, and enforcing massive, multi-tiered nested JSON schemas coming from model outputs.
- Async Pipeline Scaling: Optimizing asynchronous background tasks and workers to ingest and parse millions of multi-page scanned documents simultaneously without bottlenecking standard API performance.
Key Responsibilities
1. End-to-End Product Engineering & API Architecture
- Ship highly reliable, scalable features across the complete application stack, balancing high-performance backend logic (Python/FastAPI) with rapid frontend interfaces (React/Next.js/TypeScript).
- Design and implement flexible, low-latency APIs and robust database schemas using PostgreSQL to map complex multi-turn AI evaluation states.
- Continuously profile, optimize, and refactor existing code patterns to maximize system speed, codebase maintainability, and architectural reliability.
2. Document Automation & Pipeline Scaling
- Construct and scale secure document processing pipelines, utilizing advanced object storage systems (AWS S3) alongside rigid permission models to eliminate any risk of data leakage.
- Build advanced document extraction tools incorporating OCR and text-search mechanics to seamlessly index scanned PDFs, images, and manual expert annotations.
- Develop robust programmatic formatting engines capable of generating, rendering, and inline-editing complex nested JSON structures and multi-media document layers.
3. Systems Monitoring & Architecture Evolution
- Identify execution bottlenecks, trace performance anomalies across cloud instances, and scale distributed database storage backends to handle heavy write loads.
- Actively collaborate within a tight-knit, highly senior engineering group to break down loose product specs into clear, deterministic software systems.
- Lead the deployment of automated testing pipelines, ensuring secure handling practices and continuous delivery of platform enhancements.
Expected Outputs & Deliverables
- Production-ready backend services and matching UI views deployed to live environments weekly.
- High-performance, fully audited object storage configurations with programmatic access control policies.
- Scalable text extraction and OCR processing microservices integrated directly into the core platform pipeline.
- Documented API schemas, architecture blueprints, and clean schema validation modules committed to central repositories.
Skills Required:
- Computer / Software / It / Data
Quick Actions
Share Vacancy