Bruno Volpato's photo

Bruno Volpato

Staff Software Engineer, Applied AI & Research Infrastructure

Staff Software Engineer with 15+ years building high-scale distributed systems, query engines, and data platforms. Specializes in Applied AI and research engineering, contributing to simulation environments with AI Research partners: controlled disruptions, telemetry, and ground-truth episodes for agent evaluation and RL-adjacent training data. Also leads production query and connector systems serving hundreds of millions of queries daily.

Applied AI & Research

Generative Simulation Agent Evaluation Synthetic Data Agent Retrieval

Systems & Data

Distributed Systems Query Engines Data Platforms Production Reliability

Languages

Golang Java Python C++ SQL

Data & Platform

Kubernetes GCP Trino DataFusion Substrait Kafka

Experience

Datadog

Jan 2024 - Present

Staff Software Engineer

Applied AI & Research Infrastructure: Work with AI Research scientists and engineers on simulation environments that produce controlled disruptions, telemetry, and ground-truth investigation paths for agent evaluation and RL-adjacent training data. Contributed to the synthetic-data design and helped grow the team catalog from roughly a dozen manually created applications to thousands of generated applications.

Platform Team Leadership: Grew Query Connectors from a small group into a shared platform team. Mentored engineers across levels and conducted 50+ design and coding interviews.

Query & Data Platform: Lead query planning and connectors serving hundreds of millions of queries daily across observability use cases, spanning parsers, cost-based optimization, distributed execution, and connector reliability. Shipped Connector Server, expanding from 5 to 15 connectors and 5 to 38 catalog instances. Delivered 50x customer-facing Logs Analytics and 14x large-timeframe query speedups, with roughly 40x lower parser memory.

Open Query Systems: Became a Substrait committer (Cross-Language Serialization for Relational Algebra) and contributed across DataFusion, Trino, Iceberg, and Calcite so plans and data access are portable across engines.

Production Reliability: Serve as Core Incident Commander for severe incidents; commanded 30+ incidents and responded to 120+ in the last year, and coordinated follow-ups across monitoring, deployment gates, SLOs, and regression suites used by 30+ engineers.

Stack: Go, Java, Python, SQL, Trino, DataFusion, Apache Calcite, Kubernetes

Google

Aug 2022 - Jan 2024

Software Engineer, Tech Lead

Data Platform Leadership: Led Dataflow Platform team of 5 engineers responsible for 200+ production pipelines processing petabytes of data daily across Google Cloud.

Platform Transformation: Moved the team from bespoke customer pipeline delivery toward reusable infrastructure for teams and partners; reduced new pipeline development time by 60% and enabled 3x more use cases with the same team capacity.

Operational Reliability: Participated in 24/7 on-call handling P0/P1 incidents, customer escalations, and turnarounds while improving tests, probers, rollouts, support automation, and release safety.

Security & Delivery Hygiene: Led migration to minimal Distroless container images, eliminating over 80% of CVE surface area, and drove multi-quarter migrations and partner launches across data platform teams.

Stack: Java, Python, C++, Apache Beam, Kafka, Docker, Google Cloud Dataflow, GCP

TOTVS Labs

Sept 2015 - Jul 2022

Senior Backend Engineer

Big Data Platform: Architected scalable streaming pipelines processing billions of events daily using Kafka, Elasticsearch, and Couchbase, powering ML-driven recommendation engines for enterprise customers.

Platform Architecture: Set technical direction for backend data processing systems and helped turn high-volume event streams into reusable platform capabilities.

Performance Engineering: Built dynamic configuration-to-bytecode compilation engine, reducing rule evaluation latency by 80% and enabling real-time personalization at scale.

Stack: Java, Python, Kafka, Elasticsearch, Couchbase, Docker, JVM tuning

TOTVS

Mar 2009 - Aug 2015

Senior Software / R&D Engineer

Enterprise Integration: Designed and implemented data integration layer serving 5,000+ global manufacturing companies, processing millions of transactions daily with sub-second latency requirements.

Developer Productivity: Built source search and white-box validation tooling, improving team productivity by 30% and reducing bug escape rate.

System Architecture: Architected real-time event-driven integration pipelines for critical enterprise workflows across distributed deployments.

Stack: Java, Elasticsearch, Lucene, Spring, EJB, SQL, Oracle, Progress

Education

North Carolina State University

M.S. in Computer Science

North Carolina State University

B.S. in Computer Science