Data Engineer for AI Stack (f/m/d) @ A1 Competence Delivery Center
A1group- Location
- София
- Workplace
- —
- Employment
- —
- Salary
- —
Posted 2d ago
Strength. Care. Growth
A1 Competence Delivery Center is a vital component of A1’s telecommunications business. Acting as an expertise hub, CDC is dedicated to delivering a full range of high-quality IT, network, financial and other services to support A1’s operations across all OpCos, independent of location.
Using the power of being OneGroup and leveraging synergies, CDC enables transparency of resources, key skills and knowledge expansion and personal career growth opportunities’ enhancement, paired with job stability.
You will know we are the right place for you, if you are driven by:
- Opportunities to learn and build your career.
- Meaningful work in a stable and fast-paced company.
- Diversity of people, projects, and platforms.
- A supportive, fun, and inspiring place to work.
Would you like to join us?
Job Purpose
As a Data Engineer for AI Stack, you will make enterprise data usable by the AIaaS platform. Your focus will be data ingestion, transformation, orchestration, metadata, lineage, replay, quality, and secure integration with A1 data sources.
You will work across APIs, databases, files, object storage, event sources, and analytical platforms. You will help establish reusable integration patterns that can support AI, machine-learning, RAG, analytics, and operational automation use cases.
Role insights
- Design and implement batch, incremental, API-based, and file-based data-ingestion pipelines.
- Integrate the AIaaS platform with enterprise systems, databases, object storage, SaaS services, data lakes, and approved on-premises sources.
- Define processing patterns for transformation, enrichment, validation, deduplication, checkpointing, retries, replay, and backfill.
- Implement idempotent pipelines with clear failure and recovery semantics.
- Develop reusable connectors, Python jobs, orchestration workflows, and data-service APIs.
- Establish source-to-target mappings, schemas, contracts, and validation rules.
- Contribute to the design of operational storage, object storage, analytical layers, open-table formats, vector stores, and feature stores.
- Integrate pipelines with orchestration, metadata, lineage, observability, and data-quality services.
- Define how lineage is generated and propagated rather than relying solely on passive metadata cataloguing.
- Support RAG ingestion, document processing, embedding pipelines, feature engineering, model-training data preparation, and batch inference.
- Work with security and privacy stakeholders on data classification, minimisation, pseudonymisation, retention, and access controls.
- Provide realistic workload profiles for capacity, storage, network, and performance planning.
- Produce data-flow diagrams, interface catalogues, operational runbooks, and support documentation.
- Help distinguish reusable platform capabilities from use-case-specific engineering.
What makes you unique
- Strong data-engineering experience using Python, SQL, and modern orchestration or integration tools.
- Experience building reliable production pipelines across heterogeneous data sources.
- Knowledge of ETL/ELT, APIs, relational databases, object storage, file ingestion, and schema evolution.
- Understanding of data-quality controls, lineage, metadata, replay, backfill, and idempotency.
- Experience with workflow orchestration and automated deployment.
- Familiarity with containerised execution environments and Kubernetes-based data workloads.
- Ability to analyse failure modes and design recoverable data processes.
- Strong communication skills and the ability to work with source-system owners, architects, platform engineers, governance teams, and use-case developers.
Nice to have
- Airflow, Airbyte, Meltano, OpenMetadata, dbt, Spark, Trino, Dremio, DuckDB, Iceberg, or Delta Lake.
- MLflow, Feast, vector databases, embedding pipelines, or RAG ingestion.
- Azure, Exoscale SOS/S3, Cloudera, Synapse, Teradata, PostgreSQL, Oracle, or Microsoft SQL Server.
- Streaming technologies such as Kafka or Flink.
- Experience with OpenLineage and OpenTelemetry.
- Knowledge of AI/ML data lifecycles, feature engineering, or model-serving integration.
Our gratitude for the job done will be eternal, but we’ll also offer you:
- Innovative technologies and platforms to work with.
- Modern working environment for your comfort.
- Friendly, ambitious, and motivated teammates to support each other.
- Thousands of online and in-person learning opportunities to grow.
- Challenging assignments and career development opportunities in multinational environment.
- Attractive remuneration package.
- Flexible working schedule and opportunity for home office.
- Numerous additional goodies, including, but not limited to free A1 services, discounts, health insurance and services, sports center, childcare, team and family events, etc.
If you have any questions, please do not hesitate to contact Mariya Ivanova.
Skills
- Retrieval-Augmented Generation
- Python
- SQL
- ETL
- ELT
- Kubernetes
- Airflow
- Airbyte
- dbt
- Spark
- Trino
- Dremio
- DuckDB
- Apache Iceberg
- Delta Lake
- MLflow
- Vector Databases
- Azure
- S3
- Synapse
- PostgreSQL
- Oracle Database
- SQL Server
- Kafka
- Flink
- OpenTelemetry
More jobs at A1group
All 58Policy Operations Stream Lead (f/m/d)@ A1 Competence Delivery Center
A1group · София, бул. Черни връх 51 Б · North Macedonia · Croatia +3 · yesterday
QA Engineer (f/m/d)@ A1 Competence Delivery Center
A1group · София · 2d ago
Cyber Security Services Coordinator (f/m/d) @ A1 Competence Delivery Center
A1group · София, бул. Черни връх 51 Б · 2d ago
Senior Platform Engineer (f/m/d) @ A1 Competence Delivery Center
A1group · София · 2d ago
AI Architect & Forward Engineering (f/m/d) @ A1 Competence Delivery Center
A1group · София · 3d ago
Similar roles
Senior Analytics Engineer
Redwood Materials · San Francisco, California, United States · today
Data Analyst (Finance) Intern (6 months)
Williams-Sonoma · Singapore · today
Data Engineering Manager
Redwood Materials · McCarran, NV · today
Senior Business Intelligence Analyst
AtriCure · Mason, OH · today
Recruiting Analytics Data Engineer
Anthropic · New York City, NY · San Francisco, CA | New York City, NY | Seattle, WA · Seattle, WA · USD 285,000–380,000/yr · today
Data Analyst
UoM · Parkville, Victoria · USD 111,275–120,453/yr · today