Okx23d ago
New
New
Big Data Engineer, Web3
OtherData Operations EngineerBig Data Engineer
2 views0 saves0 applied
Quick Summary
Key Responsibilities
Design and operate large-scale distributed data systems Own the big data compute and storage infrastructure (MaxCompute/ODPS, Hologres,
Technical Tools
OtherData Operations EngineerBig Data Engineer
OKX will be prioritising applicants who have a current right to work in Singapore, and do not require OKX's sponsorship of a visa.
At OKX, we believe that the future will be reshaped by crypto, and ultimately contribute to every individual's freedom.
OKX is a leading crypto exchange, and the developer of OKX Wallet, giving millions access to crypto trading and decentralized crypto applications (dApps). OKX is also a trusted brand by hundreds of large institutions seeking access to crypto markets. We are safe and reliable, backed by our Proof of Reserves.
Across our multiple offices globally, we are united by our core principles: We Before Me, Do the Right Thing, and Get Things Done. These shared values drive our culture, shape our processes, and foster a friendly, rewarding, and diverse environment for every OK-er.
OKX is part of OKG, a group that brings the value of Blockchain to users around the world, through our leading products OKX, OKX Wallet, OKLink and more.
About the Role
~1 min readWe are building the foundational data infrastructure that powers one of the world's leading crypto exchanges. As a Big Data Platform Engineer, you will design, build, and evolve the core platform services that enable hundreds of data engineers and analysts to move fast and ship reliable pipelines. You will be at the forefront of our AI agent initiative — embedding LLM-driven capabilities directly into the platform layer so that scheduling, cost optimization, and incident response increasingly run autonomously.
Responsibilities
~2 min read- →
Platform Core: Design and operate large-scale distributed data systems
- →
Own the big data compute and storage infrastructure (MaxCompute/ODPS, Hologres, Spark)
- →
Build and maintain multi-site task orchestration that dynamically selects engines and enforces policy
- →
Drive reliability and performance improvements across batch and real-time pipelines
- →
AI Integration: Build the AI-native platform layer
- →
Develop and expose MCP (Model Context Protocol) tool interfaces so AI agents can interact with platform APIs
- →
Build the scheduling and cost-optimization agents that auto-tune resource allocation and alert severity
- →
Instrument platform telemetry to feed AI-driven SLA monitoring and anomaly detection
- →
Design context retrieval pipelines (RAG / vector search) for SQL code and config knowledge bases
- →
Tooling & DX: Evolve the developer experience
- →
Own the internal data development platform — IDE integrations, code review automation, deployment tooling
- →
Build APIs-first tools (backfill, ingestion automation) designed for future MCP integration
- →
Collaborate with data warehouse and service teams to define platform contracts
- →
Ops & Governance: Drive operational excellence
- →
Establish SLA benchmarks, cost metrics, and latency dashboards as AI optimization targets
- →
Build automated incident response and root-cause analysis pipelines
- →
Define and enforce infrastructure policies across multi-cloud environments
AI Agent Ownership — Platform Tier
- →
Scheduling Agent: auto-configure task dependencies, engine selection, cost/performance trade-offs, and alert tiers
- →
Operations Agent: detect pipeline latency, performance degradation, and schema drift; trigger remediation
- →
Incident Response Agent: trace SLA breaches to root cause, assign accountability, generate post-mortems
- →
MCP Tool Layer: design and maintain the cross-platform tool interfaces that all agents call into
-
5+ years of experience building large-scale data platforms (Hadoop/Spark/Flink or equivalent)
-
Deep expertise in distributed storage and compute systems (MaxCompute, Hologres, ClickHouse, Hive)
-
Strong software engineering skills in Java, Scala, or Python; experience with API-first design
-
Hands-on experience with task scheduling systems (Airflow, DolphinScheduler, or in-house equivalents)
-
Solid understanding of multi-cloud architectures and cost governance
-
Familiarity with LLM integration patterns: tool calling, RAG pipelines, context management
-
Experience with MCP or similar agent-tool frameworks is a strong plus
-
Passion for building systems that make other engineers 10x more productive
What We Offer
~1 min read✓Competitive total compensation package
✓L&D programs and education subsidy for employees' growth and development
✓Various team building programs and company events
✓Wellness and meal allowances
✓Comprehensive healthcare schemes for employees and dependants
✓More that we love to tell you along the process!
Location & Eligibility
Where is the job
Singapore, Singapore
On-site at the office
Who can apply
SG
Listing Details
- Posted
- July 6, 2026
- First seen
- July 6, 2026
- Last seen
- July 29, 2026
Posting Health
- Days active
- 0
- Repost count
- 0
- Trust Level
- 67%
- Scored at
- July 6, 2026
Signal breakdown
freshnesssource trustcontent trustemployer trust

Okx
greenhouse
OKX is a global cryptocurrency exchange and Web3 technology company, offering trading, wallet services, and access to decentralized finance. Founded in 2017, it serves millions of users in over 100 countries.
View company profileExternal application · ~5 min on Okx's site
Please let Okx know you found this job on Jobera.
4 other jobs at Okx
View all →Explore open roles at Okx.
Browse Similar Jobs
Newsletter
Stay ahead of the market
Get the latest job openings, salary trends, and hiring insights delivered to your inbox every week.
A
B
C
D
No spam. Unsubscribe at any time.