Srija Tummalapenta
Professional Summary
AI-Enabled Cloud Data Engineer with 5+ years of experience architecting cloud-native data platforms, scalable big data pipelines, Lakehouse architectures, and Generative AI solutions across AWS, Azure, and GCP environments. Demonstrated success processing 500GB+ of daily enterprise data through high-performance data engineering ecosystems. Expertise in PySpark, Databricks, Snowflake, Apache Spark, Kafka, RAG, and LangChain, enabling real-time analytics, AI-powered intelligence, and optimized data operations for enterprise-scale organizations.
Technical Skills
Work Experience
- Architected cloud-native PySpark ETL pipelines on Databricks and AWS EMR, processing 500GB+ of daily financial data, reducing pipeline latency and accelerating enterprise reporting workflows.
- Engineered AI-ready Lakehouse solutions using Delta Lake, AWS S3, and Medallion Architecture, enabling scalable data ingestion and transformation for 500+ business users across analytics teams.
- Resolved data quality and governance deficiencies by implementing automated validation, monitoring, and observability frameworks, reducing data incidents and strengthening regulatory compliance.
- Collaborated with engineering, analytics, risk, and compliance teams across critical data pipelines, improving data reliability, reducing compliance gaps and supporting enterprise initiatives.
- Implemented RAG and NLP solutions using LangChain, Pinecone, and LLMs to enhance contextual analytics, improving reporting accuracy and accelerating AI application delivery by 40%.
- Optimized Spark-based ETL/ELT pipelines on Databricks, PySpark, and AWS, processing large-scale retail and pharmacy data while improving data processing efficiency and platform performance.
- Developed real-time data platforms using Apache Kafka, AWS Lambda, and Snowflake to process 5,000+ transactions per second, enabling near real-time customer, prescription, and operational insights.
- Modernized legacy on-premise data warehouses through Snowflake migration and DBT-driven transformation frameworks, enhancing analytics scalability and query performance for enterprise reporting.
- Coordinated with data engineering, analytics, pharmacy operations, and business stakeholders to deliver governed data solutions that accelerated reporting and strengthened decision-making.
- Integrated AI-enabled knowledge retrieval and analytics capabilities using LangChain, LlamaIndex, LLMs, Snowflake, and AWS S3, improving information accessibility and operational efficiency.
- Spearheaded a multi-cloud data modernization initiative, migrating 5TB+ of retail sales, inventory, and supply chain data to Azure Synapse Analytics and Google BigQuery.
- Built distributed Spark and PySpark processing frameworks on Azure Databricks and Google Dataproc, processing 2TB+ of daily transactional data for retail analytics.
- Eliminated reporting latency bottlenecks by deploying real-time pipelines with Kafka, Google Pub/Sub, and Spark Streaming, reducing data delivery times from 6 hours to under 10 minutes.
- Aligned with supply chain, merchandising, and analytics stakeholders to deliver trusted data products that improved visibility into store operations and customer purchasing trends.
- Automated data governance and orchestration workflows using Apache Airflow, Cloud Composer, Azure Functions, Azure DevOps, and Jenkins, improving platform reliability across production environments.
Education
Certifications
Powered by Drivetube · Create your own profile at drivetube.ai
Explore Drivetube
- Drivetube Profile — your free digital resume — at drivetube.ai/in/your-name: one true standard resume with a Hiring Snapshot (visa status, expected salary, notice period, work preference, relocation), an ATS-ready PDF download and a single shareable link. Free forever; interview requests come from verified employers and your contact details stay masked until you accept. Documentation.
- Free Job Board — verified openings crawled ATS-by-ATS from 100,000+ real company career pages across 35 ATS platforms. Shows the true posting date from the source ATS — not when a listing was indexed — and deletes every general listing 3 days after it was actually posted. No ghost jobs, no ad-sponsored listings, no staffing reposts, no account needed. Documentation.
- Job Hunt Program — managed job hunting, a one-time purchase from $199.99. JobScout matches verified roles to your real experience band, Blend AI writes a uniquely tailored resume and cover letter for every application, and the Autofill extension fills the form — or Let Us Apply submits it for you. Documentation.
- Resume Writing Services — human-written, ATS-optimised resumes by senior career writers, from ₹499.99 / $25.99. Available in every country, written to the destination country's own standard — a US resume, UK CV, German Lebenslauf and Indian resume are genuinely different documents. A paid service, separate from the free Drivetube Profile. Documentation.
- Community Membership — from $4.99/month (₹1,999/year in India). Unlocks the gated job-board filters — visa sponsorship, security clearance, workplace and application time — plus Job-Scout AI matching, Resume Report AI, Interview AI prep sheets, Recruiter Outreach AI sent from your own Gmail, Apply or Skip triage, a daily market feed and a $10,000+ library including 23 ATS-validated resume templates. Documentation.
- Drivetube Hire — for employers — hiring with no job postings and no applications. Paste your real job description and AI matches it against candidates' true standard resumes, returning ranked candidates with a match %, matched and missing skills and written reasoning. Free tier included; employers pay, candidates never do. Documentation.
Full product documentation — every product explained, with feature-by-feature comparisons against the job boards, AI apply tools, resume services and hiring platforms people actually use.
The job board covers the United States, India, United Kingdom, Canada, Europe and Australia, and Resume Writing Services are available in every country.