Senior Software Engineer @ Wayfair

Hi, I'm Syed Ibrahim

Building large-scale data platforms, streaming systems & data lineage at scale

|

Bengaluru, Karnataka, India
About Me

Building the backbone of data infrastructure

Senior Software Engineer specializing in large-scale data platforms, streaming systems, data movement, metadata management, and data lineage. Known for combining deep technical knowledge with a practical focus on scalability, reliability, and operational excellence.

Over the years, I have built expertise across Apache Beam, Dataflow, Kafka, Dataproc, Flink, Dataplex, DataHub, CDC architectures, and cloud-native data engineering.

I have progressed from building data pipelines and platform tooling to owning critical systems and influencing broader architectural decisions. My work has contributed to significant infrastructure cost savings, platform modernization initiatives, improved reliability, and the development of data governance and lineage capabilities at scale.

$ echo "A highly ambitious data platform engineer and technology leader in the making—combining technical depth, continuous self-improvement, and a desire to create meaningful impact."|

Role

Senior Software Engineer

Focus

Data Platforms & Streaming

Company

Wayfair

5+

Years Exp

10+

Technologies

Technical Skills

Technologies I work with

Data & Streaming

Apache BeamApache KafkaApache FlinkGoogle DataflowGoogle DataprocCDC Architectures

Data Governance

DataHubGoogle DataplexData LineageMetadata ManagementData Quality FrameworksLineageScript DSL

Cloud & Infrastructure

Google Cloud PlatformKubernetesDockerCI/CD PipelinesInfrastructure as CodeCloud-Native Engineering

Languages & Tools

PythonJavaSQLYAML/JSON SpecsTerraformGit
Projects & Systems

What I've built & owned

LineageScript DSL

Designed and developed a custom Domain Specific Language (DSL) with a technical RFC for transforming raw data lineage into architectural lineage. Built collaborative YAML and JSON specifications enabling teams to define and manage lineage declaratively.

Custom DSLYAMLJSONPythonData Lineage
  • Technical RFC authored and reviewed
  • Transforms raw lineage to architectural lineage
  • Collaborative specification design

Large-Scale Data Platform

Built and owned critical systems for data movement and processing at Wayfair, contributing to significant infrastructure cost savings and platform modernization initiatives.

Apache BeamDataflowKafkaGCPPythonJava
  • Significant infrastructure cost savings
  • Platform modernization at scale
  • Improved reliability and performance

Metadata & Data Governance

Developed comprehensive metadata management and data governance capabilities, including data quality frameworks and lineage tracking across the data ecosystem.

DataHubDataplexMetadata APIsData Quality
  • End-to-end data lineage tracking
  • Automated data quality checks
  • Governance at enterprise scale

Kubernetes Queue Infrastructure

Designed scalable Kubernetes-based systems to manage custom queues, structuring environments to handle high volumes of heavy API-driven tasks with reliability and efficiency.

KubernetesDockerCustom Queue SystemAPI Infrastructure
  • High-volume task processing
  • Custom queue management
  • Scalable infrastructure design

Streaming & CDC Pipelines

Architected and implemented real-time streaming and Change Data Capture pipelines enabling near real-time data movement across distributed systems.

KafkaFlinkCDCDataprocBeam
  • Real-time data movement
  • CDC architecture design
  • Distributed system integration
Experience

Professional journey

Senior Software Engineer (Software Engineer III)

WayfairBengaluru, IndiaPresent

Owning critical data platform systems and influencing broader architectural decisions across data engineering, streaming, and governance domains.

  • Led platform modernization initiatives delivering significant infrastructure cost savings
  • Designed and built LineageScript DSL for transforming raw data lineage into architectural lineage
  • Owned critical data movement and processing systems at enterprise scale
  • Influenced architectural decisions across data engineering and governance domains
  • Built expertise across Apache Beam, Dataflow, Kafka, Flink, DataHub, and cloud-native engineering
Engineering Philosophy

How I approach engineering

Systems Thinking

Evaluating solutions through cost, performance, maintainability, scalability, and business impact.

Data-Driven Decisions

Analytical by nature, preferring informed decisions backed by data and measurable outcomes.

Operational Excellence

Combining deep technical knowledge with practical focus on reliability and scalability.

Continuous Growth

Pursuing objectives with discipline across career, technical skills, and personal development.

Get In Touch

Let's connect

Interested in discussing data engineering, distributed systems, or potential collaborations? I'd love to hear from you.