Jagadeesh Thiruveedula

GCP Data Architect | BigQuery · Cloud Modernization · Data Platform · Databricks | 500+ TiB Migrations · $2M+ Savings · 11+ yrs
Corinth, TX 76210 | +1 (945) 400-2650 | jagadeeshthiruveedula77@gmail.com | linkedin.com/in/jagadeesh-thiruveedula | Open to relocation · 50% travel

Executive Summary

GCP Data Architect with 11+ years designing enterprise data platforms and leading large-scale cloud modernization on GCP. Deep expertise in Snowflake→BigQuery, Teradata→BigQuery, Hadoop→GCP, and Couchbase→Spanner/Bigtable migrations, enterprise warehouse design, SCD2, SQL optimization, and streaming architectures. Delivered $2M+ cost savings via FinOps practices, 500+ TiB migrations, and 1B+ daily events at 99.9% uptime. Combines architecture governance, stakeholder communication, and hands-on delivery with GenAI-accelerated migration tooling to cut delivery timelines and de-risk enterprise cutovers. Customer-facing architect with discovery workshops, pre-sales, and RFP/RFI experience across healthcare, insurance, energy, and publishing sectors.

Impact Highlights

$2M+
Cost savings delivered
500+ TiB
Cloud migrations orchestrated
1B+/day
Events streamed · 99.9% uptime
200+
ETL workflows cutover · zero loss
40%
Manual refactoring cut (GenAI)
50%
Release cycles cut via IaC
30%
ETL performance uplift
10+
Enterprise programs enabled

Core Competencies

Cloud Data ArchitectureGCP (BigQuery, Dataflow, Cloud Composer, Pub/Sub, Cloud Run, Dataproc, Data Fusion), AWS, Azure, Snowflake, Databricks (Delta Lake, Unity Catalog)
Migration & ModernizationSnowflake→BigQuery, Teradata→BigQuery, Hadoop→GCP, Couchbase→Spanner/Bigtable, Mainframe/COBOL→PySpark, cross-cloud (AWS↔GCP↔Azure)
Data Modeling & WarehousingEnterprise warehouse design, dimensional modeling, SCD2, Medallion architecture (Bronze/Silver/Gold), SQL optimization, CDC patterns, data governance
Data EngineeringPySpark, Spark, Kafka, Airflow, dbt, Talend, Informatica, IICS, streaming pipelines, ETL/ELT modernization
DevOps & IaCTerraform, GitHub Actions, Bitbucket, Bamboo, CI/CD, Docker, Kubernetes, FinOps cost optimization
Security & ComplianceVPC-SC, CMEK, SSO/SAML, IAM, HIPAA / SOC 2 / FedRAMP, data lineage, access controls
GenAI AccelerationVertex AI, Azure AI Foundry, GenAI schema-mapping & code translation, RAG, LangGraph, MCP, multi-agent orchestration, MLOps, MLflow, prompt architecture, guardrails, Responsible AI governance
Languages & DBPython, SQL, Scala, Java, Shell · Oracle, MSSQL, MySQL, Teradata, Hive, Postgres, Spanner, Bigtable

Professional Experience

GCP Data Architect · John Wiley & Sons
May 2025 – Present · Remote
Consulting Engagement Embedded with Wiley platform & editorial leadership to architect and deliver a 500+ TiB Snowflake→BigQuery modernization with GenAI-accelerated tooling.
  • Snowflake → BigQuery Migration · 500+ TiB Led discovery workshops with platform and editorial leadership; architected GenAI-accelerated schema-mapping and SQL-translation pipeline cutting manual refactoring 40%; zero data loss across 200+ ETL workflows on Cloud Composer + Dataflow.
  • Enterprise Warehouse Redesign · BigQuery Redesigned dimensional model and partitioning strategy for 50M+ document corpus; implemented SCD2 patterns and SQL optimization reducing query cost and improving analyst throughput 3x.
  • Streaming & Real-Time Analytics · Pub/Sub + Dataflow Built event-driven pipelines for editorial telemetry; standardized CDC patterns and data-quality gates ensuring 99.9% uptime and audit-ready lineage.
  • GenAI Migration Accelerators Codified reusable schema-mapping, query-translation, and validation patterns into an internal platform playbook adopted by 3 downstream programs; mentored 4 engineers.
Stack: BigQuery, Dataflow, Cloud Composer, Pub/Sub, Databricks, Vertex AI, Azure AI Foundry, Terraform, GitHub Actions, Kubernetes, VPC-SC, CMEK, FinOps
GCP Data Architect · Definity
Nov 2024 – May 2025 · Remote
  • Embedded with mainframe operations; ran discovery across 12 COBOL/legacy workstreams to prioritize migration sequencing by risk and value.
  • Built GenAI-powered code-translation pipeline (Mainframe/COBOL → BigQuery SQL + PySpark) with semantic-fidelity eval harness; deployed into customer GCP VPC behind their IAM.
  • Designed high-throughput PySpark + BigQuery + Pub/Sub pipelines for batch + real-time streaming; refactored legacy logic into modern, audit-ready, anomaly-resistant patterns.
  • Mentored customer engineering team on AI-assisted modernization; codified 5 reusable transformation patterns into the internal platform playbook.
Stack: GCP, BigQuery, PySpark, Pub/Sub, GenAI Code Translation, Mainframe/COBOL, VPC, IAM
Cloud Data Architect · NRG Energy
Aug 2024 – Oct 2024 · Remote
  • Embedded with energy operations analytics; scoped cross-cloud migration with phased cutover plan for near-zero downtime.
  • Directed AWS data lake → GCP Databricks migration; built cross-cloud ingestion + transformation pipelines integrating BigQuery and Vertex AI for predictive maintenance models. Applied FinOps practices to optimize cross-cloud compute and storage spend.
  • Delivered phased cutover with <30 min downtime; unified cost analytics pipeline reduced energy-trading reporting latency by 40% and improved uptime SLA adherence to 99.95%.
Stack: AWS, GCP, Databricks, BigQuery, Vertex AI, cross-cloud ingestion
Lead Data Architect · HCA Healthcare
Sep 2022 – Jul 2024 · Nashville, TN
  • Built Generative AI accelerators automating Talend ETL + SQL → PySpark conversion — 50% cut in delivery timelines; codified as reusable framework adopted across 10+ programs.
  • Migrated 100+ TB on-prem warehouses & data lakes to GCP under HIPAA compliance; reduced infra costs and boosted ETL performance 30%.
  • Architected real-time streaming (Kafka + Pub/Sub) for 50+ data sources; ensured 100% data accuracy and audit compliance under healthcare governance.
  • Led design reviews; standardized AI-assisted coding tools; mentored 5 junior engineers to senior roles.
Stack: GCP, PySpark, Kafka, Pub/Sub, Talend, GenAI Accelerators, HIPAA, IAM, DLP
Senior Data Engineer · Charles Schwab
Apr 2019 – Aug 2022 · Austin, TX
  • Led multi-Petabyte Hadoop/Teradata → GCP migration; improved system speed & efficiency 25%; delivered $1M+ annual infra savings.
  • Rewrote/optimized ETL with PySpark + Talend processing 1B+ daily records, zero data loss; standardized CDC patterns cutting pipeline dev time 40%.
  • Built high-throughput transactional framework (PySpark + Qlik Replicate) processing 30M records/day at 99.5% SLA; automated monitoring cut data errors 15%.
  • Established IaC (Terraform + Bitbucket + Bamboo) for metadata-driven BigQuery deployments, cutting release cycles 50%; featured in Free Press Journal (May 2025) for cloud-native frameworks adopted across 10+ enterprise programs.
Stack: GCP, BigQuery, Dataproc, Data Fusion, Dataflow, PySpark, Talend, Qlik Replicate, Terraform
Data Engineer · DSO MCS Group
Aug 2018 – Mar 2019 · Plano, TX
  • Built cloud-native mortgage recovery warehousing solution integrating Mainframe, Teradata, and NAS into a unified analytics platform.
  • Developed scalable Talend Big Data streaming jobs; built reusable frameworks for file capture, SCD ingestion, snapshot ingestion, DQ checks, and Mainframe header validation adopted across the MARS team.
Stack: Talend Big Data, Hive, Java, PL/SQL, Mainframe, Teradata
Data Engineer · InnoMinds
Jun 2015 – Jul 2018 · Hyderabad, India
  • Designed and developed 20+ ETL pipelines for the CROMA warehouse using Talend and PL/SQL; built fact/dimension models processing 5M+ records/day with automated data-quality checks reducing defect rates by 25%.
  • Established reusable extraction, validation, and error-handling patterns adopted across 4 warehouse workstreams; cut new pipeline development time by 35% through framework standardization.
Stack: Talend, PL/SQL, Java, Oracle, warehouse modeling

Education

Advanced Certificate, Blockchain & Distributed Ledger Technologies
IIIT Hyderabad
2019 – 2020 · Hyderabad, India
Bachelor of Technology, Electrical Engineering
Jawaharlal Nehru Technological University
2015 · Hyderabad, India

Certifications & Languages

  • Google Cloud Professional Data Engineer (GCP PDE)
  • Talend Data Explorer
  • Spark Certified Hadoop Developer
Languages
Fluent: English, Telugu · Proficient: Hindi