AI Data Engineer

Responsibilities

AI Data Architecture & Platform Support

  • Architect and optimize high-throughput data pipelines and unstructured data processing frameworks to directly power the enterprise AI Foundation Platform.
  • Co-own the vector infrastructure alongside the AI Engineer Specialist, ensuring optimized embedding storage, indexing strategies, and fast vector search retrieval.
  • Build automated ETL/ELT pipelines capable of transforming raw enterprise data into AI-ready formats (chunking, metadata tagging, and cleaning).

Data Governance, Metadata & Cataloging

  • Implement metadata pipelines for comprehensive data catalog and data dictionary services, including schema capture, lineage tracking, and audit logging.
  • Enforce data governance, access controls, and privacy compliance across all data stores used for training, fine-tuning, and RAG retrieval.
  • Design and maintain operational data lineage to ensure transparency and auditability of data flowing into production AI models.

Advanced AI Component Integration

  • Develop and tune robust data ingestion pipelines for OCR, NLP processing, recommendation systems, and multi-modal AI inputs.
  • Optimize data tiering and caching strategies to reduce latency and infrastructure costs for live enterprise AI workloads.
  • Build secure, scalable data connectors linking legacy enterprise databases with modern LLM frameworks and autonomous agents.

Collaboration & Data Excellence

  • Partner directly with the AI Engineer Specialist to ensure seamless data delivery for real-world, high-concurrency production deployments.
  • Bridge traditional data engineering practices with modern MLOps, setting up automated data validation, quality checking, and drift detection at the data layer.
  • Provide data-level expertise during architecture reviews, ensuring scalable data design patterns across the entire AI project lifecycle.

Key Impacts

  • Ensure high-quality, AI-ready data ingestion at scale, dramatically reducing the time-to-market for enterprise AI assets.
  • Establish a secure and compliant data foundation that guarantees enterprise data privacy during AI model interactions.
  • Eliminate data bottlenecks in real-world production environments, ensuring sub-second latency for enterprise RAG and search systems.
  • Transition traditional data warehouses and lakes into future-ready, graph- and vector-enabled AI data infrastructure.

Qualification

Education & Experience

  • Master’s or Bachelor’s degree in Computer Engineering, Computer Science, Data Engineering, or a related technical field.
  • 5+ years of experience in heavy-duty Data Engineering, Big Data, or Distributed Systems, with strong production experience in an AI/ML context.
  • Proven track record of building large-scale data infrastructure supporting live, mission-critical applications.

Technical & AI Data Framework Expertise

  • Deep experience in Python and SQL, alongside modern big data tools (e.g., Spark, Kafka, or cloud equivalents).
  • Hands-on expertise in Vector Databases and unstructured data processing.
  • Strong understanding of data chunking strategies, text extraction (OCR/NLP pipelines), and metadata orchestration.
  • Familiarity with cloud data platforms (AWS/Azure/GCP) including managed data lakes, data warehouses, and cloud-native security/encryption.
  • Real-world experience implementing data catalogs, schema evolution, and automated data quality monitoring.

Leadership & Soft Skills

  • Excellent technical communication skills to align data pipeline designs with AI model engineering requirements.
  • Strong analytical and problem-solving mindset, comfortable handling noisy, unstructured enterprise data under tight deadlines.
  • Good verbal and written English communication skills for vendor engagement and cross-functional team collaboration.

เราใช้คุกกี้เพื่อมอบประสบการณ์ที่ดีในการใช้เว็บไซต์ หากคุณใช้เว็บไซต์ต่อ ถือว่าคุณยอมรับการใช้คุกกี้และ นโยบายคุกกี้ คลิก ตั้งค่า เพื่อตั้งค่าคุกกี้

ตั้งค่าความเป็นส่วนตัว

คุณสามารถเลือกการตั้งค่าคุกกี้โดยเปิด/ปิด คุกกี้ในแต่ละประเภทได้ตามความต้องการ ยกเว้น คุกกี้ที่จำเป็น

ยอมรับทั้งหมด
จัดการความเป็นส่วนตัว
  • คุกกี้ที่จำเป็น
    เปิดใช้งานตลอด

    เราใช้คุกกี้ที่จำเป็นสำหรับการทำงานพื้นฐานของเว็บไซต์ เช่น การรักษาความปลอดภัยและการบริหารจัดการเครือข่าย เป็นต้น เพื่อให้คุณสามารถเข้าชมและใช้งานเว็บไซต์ได้อย่างเป็นปกติ หากไม่มีคุกกี้นี้เว็บไซต์จะไม่สามารถทำงานได้อย่างเหมาะสม คุณไม่สามารถปิดการใช้งานคุกกี้นี้ได้
    รายละเอียดคุกกี้

  • คุกกี้เพื่อการวิเคราะห์

    คุกกี้ประเภทนี้จะทำการเก็บข้อมูลการใช้งานเว็บไซต์ของผู้เข้าชม เพื่อเป็นประโยชน์ในการวัดผล ปรับปรุง และพัฒนาประสบการณ์ที่ดีในการใช้งานเว็บไซต์ ถ้าคุณปิดการใช้งานคุกกี้นี้ เราจะไม่สามารถวัดผล ปรับปรุงและพัฒนาเว็บไซต์ได้

  • คุกกี้เพื่อปรับเนื้อหาให้เข้ากับกลุ่มเป้าหมาย

    คุกกี้ประเภทนี้จะเก็บข้อมูล เช่น เว็บไซต์หรือเนื้อหาที่คุณเยี่ยมชม เป็นต้น เพื่อให้เราสามารถนำมาวิเคราะห์ และนำเสนอเนื้อหา ให้ความเหมาะสมและตรงกับความสนใจของคุณ ถ้าคุณปิดการใช้งานคุกกี้นี้ คุณจะยังเห็นเนื้อหาและโฆษณาทั่วไป ซึ่งมีเนื้อหาและโฆษณาที่ไม่ตรงกับความสนใจของคุณ

  • คุกกี้เพื่อช่วยในการใช้งาน

    คุกกี้ประเภทนี้จะช่วยจดจำการกรอกข้อมูล เช่น สมัครงาน สนใจผลิตภัณฑ์ เพื่อให้คุณสามารถใช้งานเว็บไซต์ได้สะดวกยิ่งขึ้น โดยไม่ต้องกรอกข้อมูลใหม่ทุกครั้งที่เข้าใช้เว็บไซต์ ถ้าคุณปิดการใช้งานคุกกี้นี้ คุณอาจใช้งานเว็บไซต์ได้ไม่สะดวกและไม่เต็มประสิทธิภาพ

บันทึก