Sri Raghavendra Educational Institutions Society (R)
(Approved by AICTE, Accredited by NAAC, Affiliated to VTU, Karnataka)
Sri Krishna Institute of Technology
www.skit.org.in
Title: Traditional Business Intelligence Vs Big Data
CO addressed: CO1
Course: Big Data Analytics
Presented by: Mr. P. Kiran Kumar
Department: ISE
2
12/16/2024
/skit.org.in
(Approved by AICTE, Accredited by NAAC, Affiliated to VTU, Karnataka)
Sri Krishna Institute of Technology
Diff b/w Traditional BI & Big Data:
Traditional Business Intelligence (BI):
Analyzes structured historical business data to support reporting and decision-making.
Big Data:
Processes massive, diverse, and high-speed data to generate real-time insights and predictions.
Then add the comparison table below:
Traditional BI | Big Data |
Structured data | Structured, semi-structured & unstructured data |
Historical analysis | Historical & real-time analysis |
Data Warehouse, SQL | Hadoop, Spark, NoSQL |
Reports & Dashboards | Predictive & AI-driven analytics |
Limited scalability | Highly scalable |
GB–TB data | TB–PB+ data |
3
12/16/2024
/skit.org.in
(Approved by AICTE, Accredited by NAAC, Affiliated to VTU, Karnataka)
Sri Krishna Institute of Technology
Typical Data Warehouse Environment:
A Data Warehouse is a centralized repository that stores historical structured data collected from multiple business applications. It supports reporting, business intelligence (BI), and decision-making.
4
12/16/2024
/skit.org.in
(Approved by AICTE, Accredited by NAAC, Affiliated to VTU, Karnataka)
Sri Krishna Institute of Technology
Working of a Data Warehouse:
Step 1:
Business applications generate data.
Step 2:
ETL extracts data from different sources.
Step 3:
Data is cleaned and transformed.
Step 4:
Data is loaded into the Data Warehouse.
Step 5:
Business Intelligence tools create reports and dashboards.
5
12/16/2024
/skit.org.in
(Approved by AICTE, Accredited by NAAC, Affiliated to VTU, Karnataka)
Sri Krishna Institute of Technology
Typical Hadoop Environment:
What is Hadoop?
Hadoop is an open-source framework used to store and process Big Data across multiple computers.