Senior Data Engineer (AI-First Platform)
Software Engineering, Data Science
Latin America
USD 6k-6k / month
Job Description
- Full-Time
- Remote
- 6.000 USD
About Company
Company is an AI-native distribution company for creators and media IP. We're building a platform that takes raw content, decides what to post, where, and when, renders it, and learns from what performs — think Palantir for social media and the creator economy.
We're organized around content verticals (video, podcasts, and more to come). Each vertical plugs into the same core platform: a unified content and performance graph that gets smarter with every campaign we run. The team is small, senior, and moves fast — YC-style shipping cadence with Palantir-style rigor on systems and data.
Position Overview
We are looking for a brilliant, high-seniority Senior Data Engineer who is passionate about AI and eager to own the intersection of data science infrastructure and massive-scale data processing.
In this role, you will be responsible for architecting and orchestrating our data layers to handle extremely high volumes of unstructured data (video, images, and raw metrics) while building the infrastructure needed for our in-house AI agents to thrive.
Core Projects & Responsibilities
You will step in to own three major pillars from day one:
- In-House Social API Integration: Transition our data ingestion from third-party social media API providers into a robust, scalable, and fully managed in-house pipeline.
- High-Volume Infrastructure & Architecture: Engineer, scale, and optimize data layers and orchestration to safely store, clean, and process massive daily volumes of video, image, and raw performance data.
- LLM & Agent-Friendly Datasets: Tackle the unique technical challenge of managing and structuring natural language data (including Markdown files and unstructured layers) to serve as optimized, high-context alternative datasets for our in-house AI agents.
Requirements & Qualifications
- Experience: 5+ years of experience in Data Engineering, ideally within fast-paced startup environments or AI-centric companies.
- Language: Excellent, highly confident English communication skills. (This is a priority; you must be comfortable pitching ideas, collaborating with cross-functional teams, and documenting processes).
- Core Technical Stack: Advanced proficiency in Python and PostgreSQL is mandatory.
- Data Architecture: Strong hands-on experience with data orchestration tools, database optimization, and managing large-scale data pipelines.
- AI/LLM Curiosity: Deep interest or prior experience in figuring out how data pipelines and natural language datasets interact with Large Language Models (LLMs).
Location & Working Hours Flexibility
- UK-Based (Preferred): Hybrid model (3x a week) at our South London WeWork (near Waterloo Station). Working hours: London time with a slight late shift to overlap with international teammates.
- Latam-Based: 100% Remote. Working hours: PST (Pacific Standard Time).