WeAreDevelopers LIVE • Oct 12, 2020

Alibaba Big Data and Machine Learning Technology

Qiyang Duan

How do you process half a million transactions per second without performance degradation? Discover the bare-metal Kubernetes and machine learning architecture powering Alibaba's legendary Double 11 festival.

Pause
Mute Enter Fullscreen
#1 about 9 min

Scaling big data infrastructure for extreme transaction volumes

Handling unprecedented global e-commerce demands requires deploying highly scalable containerized architectures.

#2 about 5 min

Exploring the big data and machine learning portfolio

Managing the complete data lifecycle is achieved through an integrated ecosystem bridging storage, processing, and visualization.

#3 about 5 min

Comparing distributed processing architectures for batch and streaming

Supporting immense analytic throughput and real-time computation involves specialized engines designed for variable workload types.

#4 about 6 min

Centralizing data engineering and operations with unified IDEs

Fragmented data engineering workflows are resolved by leveraging a centralized interface for governance, scripting, and pipeline scheduling.

#5 about 4 min

Architecting machine learning projects with the PAI platform

Accelerating algorithmic deployment requires offering tools that span from visual graph builders to code-first notebook environments.

#6 about 4 min

Leveraging templates and managed notebooks for data science

Removing infrastructure overhead allows teams to rely on managed instances and prebuilt model templates for rapid experimentation.

#7 about 2 min

Engaging developers through global artificial intelligence algorithm competitions

Encouraging algorithmic innovation involves creating structured community datasets and competitions for machine learning engineers globally.

#8 about 6 min

Executing queries and scheduling pipeline jobs within DataWorks

Centralizing data operations makes writing custom integrations and automating analytical transformations significantly more efficient.

#9 about 3 min

Constructing machine learning pipelines visually via PAI Studio

Bypassing complex coding for standard statistical models is possible by linking prebuilt algorithms in a visual canvas.

#10 about 5 min

Training neural networks and switching hardware inside notebooks

Optimizing compute costs during deep learning research is simplified by dynamically toggling between CPU and GPU hardware availability.

Matching moments

1:43 min

AWS infrastructure stack and data flow pipeline overview

Artem Volk Artem Volk +1 · World Congress 2024

3:38 min

Architecting a unified data and machine learning workbench

Kapil Gupta Kapil Gupta · World Congress 2025

4:49 min

Structuring compute and data services for AI models

Radu Vunvulea Radu Vunvulea · World Congress 2025

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

1:29 min

Tech infrastructure capacity and AI product innovations

2:03 min

Solving complex engineering challenges in artificial intelligence deployment

Nico Axtmann · World Congress 2022