Database Internals Engineer
Listed on 2026-07-13
-
Software Development
Software Engineer
Our Company
At Teradata, we believe that people thrive when empowered with better information. Teradata's Autonomous Knowledge Platform activates enterprise intelligence by unifying data, knowledge, and business context to achieve tangible outcomes. With Teradata, organizations can provide agents with full context for impact when it matters. Our solution lets businesses connect and scale on premises, in the cloud, or through a hybrid approach.
Teradata delivers real business value with AI.
Join the team that built the original Teradata engine, and work on what comes next.
Inside the Office of the CTO, our Advanced Research Team is exploring a new generation of massively parallel, decentralized compute — engine architecture where parallelism is a first principle rather than a later refinement.
That choice shapes everything that follows: the structure of plans, the expression of operators, the organization of storage, and the movement of work and data across nodes. It is the kind of systems problem that can define a career, and we are approaching it with both ambition and rigor.
This is a rare opportunity to work at the frontier of database internals, learn from the people who helped define the field, and leave your mark on a new engine as it takes shape. If you want greenfield systems work backed by the people, resources, and a problem set only a company like Teradata can offer, this is the place.
What You'll DoYou will be responsible for building core components of a next-generation parallel compute engine from the ground up:
- Build components across the engine spine: SQL front end (tokenizer, parser, binder), logical/physical plan layer, rule- and cost-based optimizer, and operators (joins, aggregates, sort, scan).
- Build the storage substrate:
Arrow in-memory format, slotted-page on-disk format with checksums, buffer pool, and a B-tree or LSM access method. - Implement transactions and recovery: lock/latch management, MVCC/snapshot isolation, WAL/ARIES, checkpoints, and crash recovery.
- Add parallelism and distribution along the correct axis — exchange-based parallel execution for query work, consensus/replication/atomic-commit for data correctness — without conflating the two.
- Write design specifications before coding (schemas; null/empty/duplicate semantics; memory budget and spill behavior; cost characteristics), write tests first, implement behind the established operator interface, verify safety then speed, and report the benchmark delta.
You will work within an intentionally small Advanced Research Team inside Teradata's Office of the CTO — meaning your contributions directly shape architecture and direction rather than passing through layers of process. You will collaborate with:
- The architects and engineers behind one of the most successful massively parallel databases ever shipped, who have faced the hardest problems in this space and are now focused on what comes next.
- Senior experts who are available to challenge your thinking, share lessons learned from building the first generation, and provide mentorship on deep systems problems.
- Peers who share a commitment to rigor, quality, and first-principles engineering — a small team where everyone's work matters and is visible.
We are open to two complementary profiles for this role:
Profile A — Distributed Query Optimizer- Deep expertise in distributed query optimization: cascading optimizers, Abstract Syntax Tree (AST) binding, logical and physical plan distribution.
- Experience designing parallel execution pipelines where distribution is a first-class concern, not an afterthought.
- Familiarity with systems like Apache Data Fusion or similar distributed execution frameworks.
- Strong hands-on background with analytical/embedded engines (e.g., DuckDB, Data Fusion) and their internals.
- Deep knowledge of pipeline execution, vectorized processing, file system I/O, and open table formats (Iceberg, Delta).
- Experience with transaction management, lock managers, cache management, or OS-level scheduling.
Technical Requirements:
- Systems…
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).