Home
Big Data Training
Programming with Big Data in R Training Course

Programming with Big Data in R Training Course

The term 'Big Data' refers to solutions designed for the storage and processing of extensive data sets. Initially developed by Google, these Big Data solutions have evolved and inspired numerous similar initiatives, many of which are available as open-source projects. R is a widely adopted programming language within the financial industry.

This course is available as onsite live training in Uzbekistan or online live training.

Thank you for sending your enquiry! One of our team members will contact you shortly.

Thank you for sending your booking! One of our team members will contact you shortly.

Course Outline

Introduction to Programming Big Data with R (bpdR)

Configuring your environment to use pbdR
Overview of pbdR scope and available tools
Commonly used packages with Big Data in conjunction with pbdR

Message Passing Interface (MPI)

Utilizing pbdR MPI 5
Parallel processing techniques
Point-to-point communication
Handling Matrix sending operations
Matrix summation methods
Collective communication strategies
Matrix summation using Reduce
Scatter and Gather operations
Additional MPI communication methods

Distributed Matrices

Constructing a distributed diagonal matrix
Singular Value Decomposition (SVD) for distributed matrices
Building distributed matrices in parallel

Statistics Applications

Monte Carlo Integration
Loading datasets
Reading data across all processes
Broadcasting data from a single process
Processing partitioned data
Distributed Regression analysis
Distributed Bootstrap methods

21 Hours

Number of participants

Online

Classroom

Select Location

Please select a Venue

Price per participant

Open Training Courses require 5+ participants.

Programming with Big Data in R Training Course - Booking

Full Name *

Email *

Phone *

Job Title

Company Name

Address 1 *

City *

State / Province

Country *

Postcode *

Start Date

Tax ID

Dates are subject to availability and take place between 09:30 and 16:30.

Payment *

Bank Transfer (Invoice, PO)

Debit / Credit Card

Booking summary

Number of participants: —
Course hours: 21 Hours
Total price: —

Comments

Terms and Conditions *

I am an authorised representative of the above named client and I wish to book the above courses or services in accordance with NobleProg Terms and Conditions and Privacy Policy.

Inform me about discounts and promotions

Please read our Privacy Policy to find out how we use your data

Programming with Big Data in R Training Course - Enquiry

Full Name *

Email *

Phone *

Number of participants

Company Name

Company Address

How do you want to take the course?

Client Premises

Online

Classroom

Comments

Inform me about discounts and promotions

Please read our Privacy Policy to find out how we use your data

Programming with Big Data in R - Consultancy Enquiry

Full Name *

Phone *

Email *

Company Name

Consultancy Subject *

Consultancy Goal

Who will the consultant work with?

Consultancy Urgency *

Comments

Inform me about discounts and promotions

Please read our Privacy Policy to find out how we use your data

Testimonials (2)

The subject matter and the pace were perfect.

Tim - Ottawa Research and Development Center, Science Technology Branch, Agriculture and Agri-Food Canada

Course - Programming with Big Data in R

Michael the trainer is very knowledgeable and skillful about the subject of Big Data and R. He is very flexible and quickly customize the training meeting clients' need. He is also very capable to solve technical and subject matter problems on the go. Fantastic and professional training!.

Xiaoyuan Geng - Ottawa Research and Development Center, Science Technology Branch, Agriculture and Agri-Food Canada

Course - Programming with Big Data in R

Upcoming Courses

Programming with Big Data in R

2026-08-19 09:30

21 hours

Mirobod

33,896,260 UZS (Online)

36,896,260 UZS (Classroom)

Programming with Big Data in R

2026-09-02 09:30

21 hours

Mirobod

33,896,260 UZS (Online)

36,896,260 UZS (Classroom)

Programming with Big Data in R

2026-09-16 09:30

21 hours

Mirobod

33,896,260 UZS (Online)

36,896,260 UZS (Classroom)

Programming with Big Data in R

2026-09-30 09:30

21 hours

Mirobod

33,896,260 UZS (Online)

36,896,260 UZS (Classroom)

Related Courses

Advanced R

14 Hours

This instructor-led, live training in Uzbekistan (online or onsite) is aimed at intermediate-level advanced R users who wish to use R to build faster workflows, improve code quality, and handle more complex analysis tasks.

By the end of this training, participants will be able to: create reusable functions, improve data workflows, debug and optimize code, and produce reproducible reports.

Big Data Analytics with Google Colab and Apache Spark

14 Hours

This instructor-led, live training in Uzbekistan (online or onsite) is aimed at intermediate-level data scientists and engineers who wish to use Google Colab and Apache Spark for big data processing and analytics.

By the end of this training, participants will be able to:

Set up a big data environment using Google Colab and Spark.
Process and analyze large datasets efficiently with Apache Spark.
Visualize big data in a collaborative environment.
Integrate Apache Spark with cloud-based tools.

Introductory R (Basic to Intermediate)

14 Hours

This instructor-led live training in Uzbekistan (online or onsite) is aimed at beginner-level data analysts who wish to use R programming to manipulate data, perform basic data analysis, and create compelling visualizations for insights.

By the end of this training, participants will be able to:

Understand the basics of R Programming.
Apply fundamental data science processes.
Create visual representations of data.

Data and Analytics - from the ground up

42 Hours

In today's business landscape, data analytics is an indispensable tool. This course prioritizes the development of practical, hands-on skills for effective data analysis. Its primary goal is to empower participants to provide evidence-based answers to critical questions:

What has happened?

processing and analyzing data
producing informative data visualizations

What will happen?

forecasting future performance
evaluating forecasts

What should happen?

turning data into evidence-based business decisions
optimizing processes

Data Analysis with Python, R, Power Query, and Power BI

21 Hours

This instructor-led, live training in Uzbekistan (online or onsite) is aimed at beginner-level professionals who wish to clean and analyze data, make statistical projections, and create insightful visualizations using these tools.

By the end of this training, participants will be able to:

Understand the basics of Python, R, Power Query, and Power BI for data analysis.
Clean and organize datasets using Python and Power Query.
Perform statistical analysis and projections with R.
Create professional dashboards and reports with Power BI.
Integrate and analyze data from multiple sources effectively.

Apache NiFi for Administrators

21 Hours

Apache NiFi is an open-source, flow-based data integration and event-processing platform. It enables automated, real-time data routing, transformation, and system mediation between disparate systems, with a web-based UI and fine-grained control.

This instructor-led, live training (onsite or remote) is aimed at intermediate-level administrators and engineers who wish to deploy, manage, secure, and optimize NiFi dataflows in production environments.

By the end of this training, participants will be able to:

Install, configure, and maintain Apache NiFi clusters.
Design and manage dataflows from varied sources and sinks.
Implement flow automation, routing, and transformation logic.
Optimize performance, monitor operations, and troubleshoot issues.

Format of the Course

Interactive lecture with real-world architecture discussion.
Hands-on labs: building, deploying, and managing flows.
Scenario-based exercises in a live-lab environment.

Course Customization Options

To request a customized training for this course, please contact us to arrange.

Python and Spark for Big Data for Banking (PySpark)

14 Hours

Renowned for its clear syntax and code readability, Python is a high-level programming language. Spark serves as a powerful engine for processing big data, enabling efficient querying, analysis, and transformation. PySpark bridges the two, allowing users to interface Spark with Python.

Target Audience: This course is designed for intermediate-level banking professionals who are already familiar with Python and Spark and wish to enhance their expertise in big data processing and machine learning.

PySpark and Machine Learning

21 Hours

This training offers a hands-on introduction to developing scalable data processing and Machine Learning workflows with PySpark. Participants will learn how Apache Spark functions within contemporary Big Data ecosystems and how to efficiently manage large datasets using distributed computing principles.

Apache Spark Fundamentals

21 Hours

This instructor-led, live training in Uzbekistan (online or onsite) is aimed at engineers who wish to set up and deploy Apache Spark system for processing very large amounts of data.

By the end of this training, participants will be able to:

Install and configure Apache Spark.
Quickly process and analyze very large data sets.
Understand the difference between Apache Spark and Hadoop MapReduce and when to use which.
Integrate Apache Spark with other machine learning tools.

Administration of Apache Spark

35 Hours

This instructor-led, live training in Uzbekistan (online or onsite) is aimed at beginner-level to intermediate-level system administrators who wish to deploy, maintain, and optimize Spark clusters.

By the end of this training, participants will be able to:

Install and configure Apache Spark in various environments.
Manage cluster resources and monitor Spark applications.
Optimize the performance of Spark clusters.
Implement security measures and ensure high availability.
Debug and troubleshoot common Spark issues.

Apache Spark in the Cloud

21 Hours

The initial learning curve for Apache Spark can be steep, requiring significant effort before yielding results. This course is designed to help you navigate that challenging early stage. Upon completion, participants will grasp the fundamentals of Apache Spark, clearly distinguish between RDDs and DataFrames, and gain proficiency in both Python and Scala APIs. Learners will also develop a solid understanding of executors, tasks, and other core concepts. Aligned with industry best practices, the course places strong emphasis on cloud deployment strategies, specifically within Databricks and AWS environments. Students will also explore the distinctions between AWS EMR and AWS Glue, one of AWS’s most recent Spark services.

AUDIENCE:

Data Engineers, DevOps Professionals, Data Scientists

Python and Spark for Big Data (PySpark)

21 Hours

In this instructor-led live training in Uzbekistan, participants will learn how to leverage Python and Spark together to analyze big data while completing hands-on exercises.

By the end of this training, participants will be able to:

Master the use of Spark with Python to analyze Big Data.
Complete exercises that simulate real-world scenarios.
Apply various tools and techniques for big data analysis using PySpark.

Python, Spark, and Hadoop for Big Data

21 Hours

This instructor-led live training in Uzbekistan (available online or on-site) is designed for developers who want to use and integrate Spark, Hadoop, and Python to process, analyze, and transform large and complex data sets.

By the end of this training, participants will be able to:

Set up the necessary environment to start processing big data with Spark, Hadoop, and Python.
Understand the features, core components, and architecture of Spark and Hadoop.
Learn how to integrate Spark, Hadoop, and Python for big data processing.
Explore the tools in the Spark ecosystem (Spark MlLib, Spark Streaming, Kafka, Sqoop, Kafka, and Flume).
Build collaborative filtering recommendation systems similar to Netflix, YouTube, Amazon, Spotify, and Google.
Use Apache Mahout to scale machine learning algorithms.

Stratio: Rocket and Intelligence Modules with PySpark

14 Hours

Stratio is a data-centric platform that integrates big data, AI, and governance into a single solution. Its Rocket and Intelligence modules enable rapid data exploration, transformation, and advanced analytics in enterprise environments.

This instructor-led, live training (online or onsite) is aimed at intermediate-level data professionals who wish to use the Rocket and Intelligence modules in Stratio effectively with PySpark, focusing on looping structures, user-defined functions, and advanced data logic.

By the end of this training, participants will be able to:

Navigate and work within the Stratio platform using Rocket and Intelligence modules.
Apply PySpark in the context of data ingestion, transformation, and analysis.
Use loops and conditional logic to control data workflows and feature engineering tasks.
Create and manage user-defined functions (UDFs) for reusable data operations in PySpark.

Format of the Course

Interactive lecture and discussion.
Lots of exercises and practice.
Hands-on implementation in a live-lab environment.

Course Customization Options

To request a customized training for this course, please contact us to arrange.

Programming with Big Data in R Training Course

Course Outline

Introduction to Programming Big Data with R (bpdR)

Message Passing Interface (MPI)

Distributed Matrices

Statistics Applications

Testimonials (2)

Tim - Ottawa Research and Development Center, Science Technology Branch, Agriculture and Agri-Food Canada

Course - Programming with Big Data in R

Xiaoyuan Geng - Ottawa Research and Development Center, Science Technology Branch, Agriculture and Agri-Food Canada

Course - Programming with Big Data in R

Upcoming Courses

Programming with Big Data in R

Programming with Big Data in R

Programming with Big Data in R

Programming with Big Data in R

Related Categories

This site in other countries/regions

Europe

Asia Pacific

North America

South America

Africa / Middle East

Other sites

Programming with Big Data in R Training Course

Course Outline

Introduction to Programming Big Data with R (bpdR)

Message Passing Interface (MPI)

Distributed Matrices

Statistics Applications

Testimonials (2)

Tim - Ottawa Research and Development Center, Science Technology Branch, Agriculture and Agri-Food Canada

Course - Programming with Big Data in R

Xiaoyuan Geng - Ottawa Research and Development Center, Science Technology Branch, Agriculture and Agri-Food Canada

Course - Programming with Big Data in R

Upcoming Courses

Programming with Big Data in R

Programming with Big Data in R

Programming with Big Data in R

Programming with Big Data in R

Related Courses

Advanced R

Big Data Analytics with Google Colab and Apache Spark

Introductory R (Basic to Intermediate)

Data and Analytics - from the ground up

What has happened?

What will happen?

What should happen?

Data Analysis with Python, R, Power Query, and Power BI

Apache NiFi for Administrators

Python and Spark for Big Data for Banking (PySpark)

PySpark and Machine Learning

Apache Spark Fundamentals

Administration of Apache Spark

Apache Spark in the Cloud

Python and Spark for Big Data (PySpark)

Python, Spark, and Hadoop for Big Data

Stratio: Rocket and Intelligence Modules with PySpark

Related Categories

Big Data

R Language

This site in other countries/regions

Europe

Asia Pacific

North America

South America

Africa / Middle East

Other sites