Sách Python for Data Analysis - 3rd Edition của Wes McKinney

Khám phá Oceanofpdf com với cuốn sách Python for Data Analysis, ấn bản thứ 3 của Wes McKinney, hướng dẫn phân tích dữ liệu hiệu quả.

Trường đại học

O'Reilly Media

Chuyên ngành

Data Analysis

Người đăng

Ẩn danh

Thể loại

book

2022

582
2
0

Phí lưu trữ

135 Point

Mục lục chi tiết

Preface

1. CHƯƠNG 1: What Is This Book About?

1.1. What Kinds of Data?

1.2. Why Python for Data Analysis?

1.3. Essential Python Libraries

1.4. Installation and Setup

1.5. Community and Conferences

1.6. Navigating This Book

2. CHƯƠNG 2: Python Language Basics, IPython, and Jupyter Notebooks

2.1. The Python Interpreter

2.2. IPython Basics

2.3. Python Language Basics

3. CHƯƠNG 3: Built-In Data Structures, Functions, and Files

3.1. Data Structures and Sequences

3.2. Functions

3.3. Files and the Operating System

4. CHƯƠNG 4: NumPy Basics: Arrays and Vectorized Computation

4.1. The NumPy ndarray: A Multidimensional Array Object

4.2. Pseudorandom Number Generation

4.3. Universal Functions: Fast Element-Wise Array Functions

4.4. Array-Oriented Programming with Arrays

4.5. File Input and Output with Arrays

4.7. Example: Random Walks

5. CHƯƠNG 5: Getting Started with pandas

5.1. Introduction to pandas Data Structures

5.2. Essential Functionality

5.3. Summarizing and Computing Descriptive Statistics

6. CHƯƠNG 6: Data Loading, Storage, and File Formats

6.1. Reading and Writing Data in Text Format

6.2. Binary Data Formats

6.3. Interacting with Web APIs

6.4. Interacting with Databases

7. CHƯƠNG 7: Data Cleaning and Preparation

7.1. Handling Missing Data

7.2. Data Transformation

7.3. Extension Data Types

7.4. String Manipulation

7.5. Categorical Data

8. CHƯƠNG 8: Data Wrangling: Join, Combine, and Reshape

8.1. Hierarchical Indexing

8.2. Combining and Merging Datasets

8.3. Reshaping and Pivoting

9. CHƯƠNG 9: Plotting and Visualization

9.1. A Brief matplotlib API Primer

9.2. Plotting with pandas and seaborn

9.3. Other Python Visualization Tools

10. CHƯƠNG 10: Data Aggregation and Group Operations

10.1. How to Think About Group Operations

10.2. Data Aggregation

10.3. Apply: General split-apply-combine

10.4. Group Transforms and “Unwrapped” GroupBys

10.5. Pivot Tables and Cross-Tabulation

11. CHƯƠNG 11: Date and Time Data Types and Tools

11.1. Date and Time Data Types and Tools

11.2. Time Series Basics

11.3. Date Ranges, Frequencies, and Shifting

11.4. Time Zone Handling

11.5. Periods and Period Arithmetic

11.6. Resampling and Frequency Conversion

11.7. Moving Window Functions

12. CHƯƠNG 12: Introduction to Modeling Libraries in Python

12.1. Interfacing Between pandas and Model Code

12.2. Creating Model Descriptions with Patsy

12.3. Introduction to statsmodels

12.4. Introduction to scikit-learn

13. CHƯƠNG 13: Data Analysis Examples

13.1. ndarray Object Internals

13.2. MovieLens 1M Dataset

13.3. US Baby Names 1880–2010

13.4. USDA Food Database

13.5. 2012 Federal Election Commission Database

Appendices

A. Advanced Array Manipulation

A.1. Reshaping Arrays

A.2. Broadcasting

A.3. Advanced ufunc Usage

A.4. Structured and Record Arrays

A.5. More About Sorting

A.6. Writing Fast NumPy Functions with Numba

A.7. Advanced Array Input and Output

A.8. Performance Tips

B. More on the IPython System

B.1. Terminal Keyboard Shortcuts

B.2. About Magic Commands

B.3. Using the Command History

B.4. Interacting with the Operating System

B.5. Software Development Tools

B.6. Tips for Productive Code Development Using IPython

B.7. Advanced IPython Features

Tóm tắt

I. Hướng Dẫn Tổng Quan Về Phân Tích Dữ Liệu Với Python

Phân tích dữ liệu là một lĩnh vực quan trọng trong khoa học dữ liệu. Sách 'Python for Data Analysis' của Wes McKinney cung cấp hướng dẫn chi tiết về cách sử dụng Python để xử lý và phân tích dữ liệu. Tài liệu này không chỉ giúp người đọc hiểu rõ về PandasNumPy, mà còn cung cấp các ví dụ thực tiễn để giải quyết các vấn đề phân tích dữ liệu phức tạp.

1.1. Giới Thiệu Về Wes McKinney Và Tác Phẩm Của Ông

Wes McKinney là người sáng lập dự án Pandas. Ông đã viết cuốn sách này để giúp các nhà phân tích dữ liệu nắm vững các công cụ cần thiết cho việc xử lý dữ liệu. Cuốn sách này là tài liệu tham khảo quan trọng cho những ai muốn tìm hiểu về phân tích dữ liệu với Python.

1.2. Tại Sao Nên Chọn Python Cho Phân Tích Dữ Liệu

Python là ngôn ngữ lập trình phổ biến trong lĩnh vực phân tích dữ liệu nhờ vào cú pháp dễ hiểu và thư viện phong phú. Sự linh hoạt của Python cho phép người dùng thực hiện nhiều tác vụ khác nhau từ xử lý dữ liệu đến machine learning.

II. Những Thách Thức Trong Phân Tích Dữ Liệu Với Python

Mặc dù Python là một công cụ mạnh mẽ, nhưng việc phân tích dữ liệu vẫn gặp nhiều thách thức. Các vấn đề như dữ liệu không đầy đủ, dữ liệu không đồng nhất và việc lựa chọn phương pháp phân tích phù hợp là những khó khăn thường gặp. Cuốn sách của Wes McKinney giúp người đọc nhận diện và giải quyết những vấn đề này.

2.1. Vấn Đề Dữ Liệu Thiếu

Dữ liệu thiếu có thể gây ra những sai lệch trong kết quả phân tích. Wes McKinney cung cấp các phương pháp để xử lý dữ liệu thiếu, bao gồm việc loại bỏ hoặc thay thế các giá trị thiếu.

2.2. Dữ Liệu Không Đồng Nhất

Dữ liệu không đồng nhất có thể đến từ nhiều nguồn khác nhau. Việc chuẩn hóa dữ liệu là cần thiết để đảm bảo tính chính xác trong phân tích. Cuốn sách hướng dẫn cách sử dụng Pandas để chuẩn hóa dữ liệu hiệu quả.

III. Phương Pháp Phân Tích Dữ Liệu Với Python

Cuốn sách 'Python for Data Analysis' giới thiệu nhiều phương pháp phân tích dữ liệu khác nhau. Các phương pháp này bao gồm xử lý dữ liệu, trực quan hóa dữ liệu, và phân tích thống kê. Mỗi phương pháp đều có những công cụ và kỹ thuật riêng để đạt được kết quả tốt nhất.

3.1. Xử Lý Dữ Liệu Với Pandas

Pandas là thư viện chính để xử lý dữ liệu trong Python. Nó cung cấp các công cụ mạnh mẽ để làm sạch, biến đổi, và phân tích dữ liệu. Các chức năng như groupBymerge giúp người dùng dễ dàng thao tác với dữ liệu.

3.2. Trực Quan Hóa Dữ Liệu Với Matplotlib

Trực quan hóa dữ liệu là một phần quan trọng trong phân tích. Matplotlib là thư viện phổ biến để tạo ra các biểu đồ và đồ thị. Cuốn sách hướng dẫn cách sử dụng Matplotlib để tạo ra các hình ảnh trực quan giúp người đọc dễ dàng hiểu dữ liệu.

IV. Ứng Dụng Thực Tiễn Của Phân Tích Dữ Liệu

Phân tích dữ liệu có nhiều ứng dụng thực tiễn trong các lĩnh vực như tài chính, y tế, và marketing. Cuốn sách của Wes McKinney cung cấp các ví dụ thực tế để minh họa cách áp dụng các kỹ thuật phân tích dữ liệu vào các tình huống cụ thể.

4.1. Phân Tích Dữ Liệu Trong Tài Chính

Trong lĩnh vực tài chính, phân tích dữ liệu giúp các nhà đầu tư đưa ra quyết định thông minh hơn. Cuốn sách cung cấp các ví dụ về cách sử dụng Python để phân tích dữ liệu tài chính và dự đoán xu hướng thị trường.

4.2. Phân Tích Dữ Liệu Trong Y Tế

Phân tích dữ liệu trong y tế có thể giúp cải thiện chất lượng chăm sóc sức khỏe. Các ví dụ trong sách cho thấy cách sử dụng Python để phân tích dữ liệu bệnh nhân và phát hiện các xu hướng sức khỏe.

V. Kết Luận Và Tương Lai Của Phân Tích Dữ Liệu Với Python

Phân tích dữ liệu với Python đang ngày càng trở nên quan trọng trong thế giới hiện đại. Cuốn sách của Wes McKinney không chỉ cung cấp kiến thức cơ bản mà còn mở ra hướng đi cho tương lai của phân tích dữ liệu. Sự phát triển của các công cụ và thư viện mới sẽ tiếp tục thúc đẩy lĩnh vực này.

5.1. Tương Lai Của Python Trong Phân Tích Dữ Liệu

Python sẽ tiếp tục là ngôn ngữ chính trong phân tích dữ liệu nhờ vào sự phát triển không ngừng của các thư viện như PandasNumPy. Sự hỗ trợ từ cộng đồng cũng sẽ giúp Python duy trì vị thế của mình.

5.2. Xu Hướng Mới Trong Phân Tích Dữ Liệu

Các xu hướng như machine learningtrí tuệ nhân tạo đang ngày càng được tích hợp vào phân tích dữ liệu. Cuốn sách cung cấp cái nhìn sâu sắc về cách các công nghệ này có thể được áp dụng trong thực tiễn.

15/07/2025
Oceanofpdf com python for data analysis 3rd edition wes mckinney

Trích đoạn nội dung tài liệu

Python for Data Analysis Data Wrangling with pandas, NumPy & Jupyter by Wes McKinney Python for Data Analysis Get the definitive handbook for manipulating, processing, cleaning, and crunching datasets in Python. Updated for “With this new edition, Python 3.4, the third edition of this hands- Wes has updated his on guide is packed with practical case studies that show you book to ensure it remains how to solve a broad set of data analysis problems effectively. the go-to resource for You’ll learn the latest versions of pandas, NumPy, and Jupyter in the process. all things related to data analysis with Python Written by Wes McKinney, the creator of the Python pandas project, this book is a practical, modern introduction to and pandas.

I cannot data science tools in Python. It’s ideal for analysts new to recommend this book Python and for Python programmers new to data science highly enough.” and scientific computing. Data files and related material are —Paul Barry available on GitHub. Lecturer and author of O’Reilly’s Head First Python • Use the Jupyter notebook and the IPython shell for exploratory computing Wes McKinney, cofounder and chief • Learn basic and advanced features in NumPy technology officer of Voltron Data, is • Get started with data analysis tools in the pandas library an active member of the Python data community and an advocate for Python • Use flexible tools to load, clean, transform, merge, and use in data analysis, finance, and reshape data statistical computing applications.

A • Create informative visualizations with matplotlib graduate of MIT, he’s also a member of the project management committees • Apply the pandas groupBy facility to slice, dice, and for the Apache Software Foundation’s summarize datasets Apache Arrow and Apache Parquet projects. • Analyze and manipulate regular and irregular time series data • Learn how to solve real-world data analysis problems with thorough, detailed examples DATA Twitter: @oreillymedia linkedin.com/company/oreilly-media US $69.com/oreillymedia ISBN: 978-1-098-10403-0 56999 9 781098 104030 THIRD EDITION Python for Data Analysis Data Wrangling with pandas, NumPy, and Jupyter Wes McKinney Beijing Boston Farnham Sebastopol Tokyo Python for Data Analysis by Wes McKinney Copyright © 2022 Wesley McKinney. All rights reserved. Printed in the United States of America.

Published by O’Reilly Media, Inc., 1005 Gravenstein Highway North, Sebastopol, CA 95472. O’Reilly books may be purchased for educational, business, or sales promotional use. Online editions are also available for most titles (http://oreilly. For more information, contact our corporate/institutional sales department: 800-998-9938 or corporate@oreilly.

Acquisitions Editor: Jessica Haberman Indexer: Sue Klefstad Development Editor: Angela Rufino Interior Designer: David Futato Production Editor: Christopher Faucher Cover Designer: Karen Montgomery Copyeditor: Sonia Saruba Illustrator: Kate Dullea Proofreader: Piper Editorial Consulting, LLC October 2012: First Edition October 2017: Second Edition August 2022: Third Edition Revision History for the Third Edition 2022-08-12: First Release See https://www.com/catalog/errata.csp?isbn=0636920519829 for release details. The O’Reilly logo is a registered trademark of O’Reilly Media, Inc. Python for Data Analysis, the cover image, and related trade dress are trademarks of O’Reilly Media, Inc. While the publisher and the author have used good faith efforts to ensure that the information and instructions contained in this work are accurate, the publisher and the author disclaim all responsibility for errors or omissions, including without limitation responsibility for damages resulting from the use of or reliance on this work.

Use of the information and instructions contained in this work is at your own risk. If any code samples or other technology this work contains or describes is subject to open source licenses or the intellectual property rights of others, it is your responsibility to ensure that your use thereof complies with such licenses and/or rights. 978-1-098-10403-0 [LSI] Table of Contents Preface.1 What Is This Book About? 1 What Kinds of Data? 1 1.2 Why Python for Data Analysis? 2 Python as Glue 3 Solving the “Two-Language” Problem 3 Why Not Python? 3 1.3 Essential Python Libraries 4 NumPy 4 pandas 5 matplotlib 6 IPython and Jupyter 6 SciPy 7 scikit-learn 8 statsmodels 8 Other Packages 9 1.4 Installation and Setup 9 Miniconda on Windows 9 GNU/Linux 10 Miniconda on macOS 11 Installing Necessary Packages 11 Integrated Development Environments and Text Editors 12 1.5 Community and Conferences 13 1.6 Navigating This Book 14 Code Examples 15 iii Data for Examples 15 Import Conventions 16 2. Python Language Basics, IPython, and Jupyter Notebooks.1 The Python Interpreter 18 2.2 IPython Basics 19 Running the IPython Shell 19 Running the Jupyter Notebook 20 Tab Completion 23 Introspection 25 2.3 Python Language Basics 26 Language Semantics 26 Scalar Types 34 Control Flow 42 2.

Built-In Data Structures, Functions, and Files.1 Data Structures and Sequences 47 Tuple 47 List 51 Dictionary 55 Set 59 Built-In Sequence Functions 62 List, Set, and Dictionary Comprehensions 63 3.2 Functions 65 Namespaces, Scope, and Local Functions 67 Returning Multiple Values 68 Functions Are Objects 69 Anonymous (Lambda) Functions 70 Generators 71 Errors and Exception Handling 74 3.3 Files and the Operating System 76 Bytes and Unicode with Files 80 3. NumPy Basics: Arrays and Vectorized Computation.1 The NumPy ndarray: A Multidimensional Array Object 85 Creating ndarrays 86 Data Types for ndarrays 88 Arithmetic with NumPy Arrays 91 Basic Indexing and Slicing 92 iv | Table of Contents Boolean Indexing 97 Fancy Indexing 100 Transposing Arrays and Swapping Axes 102 4.2 Pseudorandom Number Generation 103 4.3 Universal Functions: Fast Element-Wise Array Functions 105 4.4 Array-Oriented Programming with Arrays 108 Expressing Conditional Logic as Array Operations 110 Mathematical and Statistical Methods 111 Methods for Boolean Arrays 113 Sorting 114 Unique and Other Set Logic 115 4.5 File Input and Output with Arrays 116 4.7 Example: Random Walks 118 Simulating Many Random Walks at Once 120 4. Getting Started with pandas.1 Introduction to pandas Data Structures 124 Series 124 DataFrame 129 Index Objects 136 5.2 Essential Functionality 138 Reindexing 138 Dropping Entries from an Axis 141 Indexing, Selection, and Filtering 142 Arithmetic and Data Alignment 152 Function Application and Mapping 158 Sorting and Ranking 160 Axis Indexes with Duplicate Labels 164 5.3 Summarizing and Computing Descriptive Statistics 165 Correlation and Covariance 168 Unique Values, Value Counts, and Membership 170 5. Data Loading, Storage, and File Formats.1 Reading and Writing Data in Text Format 175 Reading Text Files in Pieces 182 Writing Data to Text Format 184 Working with Other Delimited Formats 185 JSON Data 187 Table of Contents | v XML and HTML: Web Scraping 189 6.2 Binary Data Formats 193 Reading Microsoft Excel Files 194 Using HDF5 Format 195 6.3 Interacting with Web APIs 197 6.4 Interacting with Databases 199 6.

Data Cleaning and Preparation.1 Handling Missing Data 203 Filtering Out Missing Data 205 Filling In Missing Data 207 7.2 Data Transformation 209 Removing Duplicates 209 Transforming Data Using a Function or Mapping 211 Replacing Values 212 Renaming Axis Indexes 214 Discretization and Binning 215 Detecting and Filtering Outliers 217 Permutation and Random Sampling 219 Computing Indicator/Dummy Variables 221 7.3 Extension Data Types 224 7.4 String Manipulation 227 Python Built-In String Object Methods 227 Regular Expressions 229 String Functions in pandas 232 7.5 Categorical Data 235 Background and Motivation 236 Categorical Extension Type in pandas 237 Computations with Categoricals 240 Categorical Methods 242 7. Data Wrangling: Join, Combine, and Reshape.1 Hierarchical Indexing 247 Reordering and Sorting Levels 250 Summary Statistics by Level 251 Indexing with a DataFrame’s columns 252 8.2 Combining and Merging Datasets 253 Database-Style DataFrame Joins 254 Merging on Index 259 vi | Table of Contents Concatenating Along an Axis 263 Combining Data with Overlap 268 8.3 Reshaping and Pivoting 270 Reshaping with Hierarchical Indexing 270 Pivoting “Long” to “Wide” Format 273 Pivoting “Wide” to “Long” Format 277 8. Plotting and Visualization.1 A Brief matplotlib API Primer 282 Figures and Subplots 283 Colors, Markers, and Line Styles 288 Ticks, Labels, and Legends 290 Annotations and Drawing on a Subplot 294 Saving Plots to File 296 matplotlib Configuration 297 9.2 Plotting with pandas and seaborn 298 Line Plots 298 Bar Plots 301 Histograms and Density Plots 309 Scatter or Point Plots 311 Facet Grids and Categorical Data 314 9.3 Other Python Visualization Tools 317 9. Data Aggregation and Group Operations.1 How to Think About Group Operations 320 Iterating over Groups 324 Selecting a Column or Subset of Columns 326 Grouping with Dictionaries and Series 327 Grouping with Functions 328 Grouping by Index Levels 328 10.2 Data Aggregation 329 Column-Wise and Multiple Function Application 331 Returning Aggregated Data Without Row Indexes 335 10.3 Apply: General split-apply-combine 335 Suppressing the Group Keys 338 Quantile and Bucket Analysis 338 Example: Filling Missing Values with Group-Specific Values 340 Example: Random Sampling and Permutation 343 Example: Group Weighted Average and Correlation 344 Table of Contents | vii Example: Group-Wise Linear Regression 347 10.4 Group Transforms and “Unwrapped” GroupBys 347 10.5 Pivot Tables and Cross-Tabulation 351 Cross-Tabulations: Crosstab 354 10.1 Date and Time Data Types and Tools 358 Converting Between String and Datetime 359 11.2 Time Series Basics 361 Indexing, Selection, Subsetting 363 Time Series with Duplicate Indices 365 11.3 Date Ranges, Frequencies, and Shifting 366 Generating Date Ranges 367 Frequencies and Date Offsets 370 Shifting (Leading and Lagging) Data 371 11.4 Time Zone Handling 374 Time Zone Localization and Conversion 375 Operations with Time Zone-Aware Timestamp Objects 377 Operations Between Different Time Zones 378 11.5 Periods and Period Arithmetic 379 Period Frequency Conversion 380 Quarterly Period Frequencies 382 Converting Timestamps to Periods (and Back) 384 Creating a PeriodIndex from Arrays 385 11.6 Resampling and Frequency Conversion 387 Downsampling 388 Upsampling and Interpolation 391 Resampling with Periods 392 Grouped Time Resampling 394 11.7 Moving Window Functions 396 Exponentially Weighted Functions 399 Binary Moving Window Functions 401 User-Defined Moving Window Functions 402 11.

Introduction to Modeling Libraries in Python.1 Interfacing Between pandas and Model Code 405 12.2 Creating Model Descriptions with Patsy 408 Data Transformations in Patsy Formulas 410 Categorical Data and Patsy 412 viii | Table of Contents 12.3 Introduction to statsmodels 415 Estimating Linear Models 415 Estimating Time Series Processes 419 12.4 Introduction to scikit-learn 420 12. Data Analysis Examples.1 Bitly Data from 1.gov 425 Counting Time Zones in Pure Python 426 Counting Time Zones with pandas 428 13.2 MovieLens 1M Dataset 435 Measuring Rating Disagreement 439 13.3 US Baby Names 1880–2010 443 Analyzing Naming Trends 448 13.4 USDA Food Database 457 13.5 2012 Federal Election Commission Database 463 Donation Statistics by Occupation and Employer 466 Bucketing Donation Amounts 469 Donation Statistics by State 471 13.1 ndarray Object Internals 473 NumPy Data Type Hierarchy 474 A.2 Advanced Array Manipulation 476 Reshaping Arrays 476 C Versus FORTRAN Order 478 Concatenating and Splitting Arrays 479 Repeating Elements: tile and repeat 481 Fancy Indexing Equivalents: take and put 483 A.3 Broadcasting 484 Broadcasting over Other Axes 487 Setting Array Values by Broadcasting 489 A.4 Advanced ufunc Usage 490 ufunc Instance Methods 490 Writing New ufuncs in Python 493 A.5 Structured and Record Arrays 493 Nested Data Types and Multidimensional Fields 494 Why Use Structured Arrays? 495 A.6 More About Sorting 495 Indirect Sorts: argsort and lexsort 497 Table of Contents | ix Alternative Sort Algorithms 498 Partially Sorting Arrays 499 numpy.searchsorted: Finding Elements in a Sorted Array 500 A.7 Writing Fast NumPy Functions with Numba 501 Creating Custom numpy.ufunc Objects with Numba 502 A.8 Advanced Array Input and Output 503 Memory-Mapped Files 503 HDF5 and Other Array Storage Options 504 A.9 Performance Tips 505 The Importance of Contiguous Memory 505 B. More on the IPython System.1 Terminal Keyboard Shortcuts 509 B.2 About Magic Commands 510 The %run Command 512 Executing Code from the Clipboard 513 B.3 Using the Command History 514 Searching and Reusing the Command History 514 Input and Output Variables 515 B.4 Interacting with the Operating System 516 Shell Commands and Aliases 517 Directory Bookmark System 518 B.5 Software Development Tools 519 Interactive Debugger 519 Timing Code: %time and %timeit 523 Basic Profiling: %prun and %run -p 525 Profiling a Function Line by Line 527 B.6 Tips for Productive Code Development Using IPython 529 Reloading Module Dependencies 529 Code Design Tips 530 B.

Nội dung được bảo vệ bản quyền — Tải xuống đầy đủ

Tài liệu "Hướng Dẫn Phân Tích Dữ Liệu Với Python" của Wes McKinney cung cấp một cái nhìn sâu sắc về cách sử dụng Python để phân tích và xử lý dữ liệu hiệu quả. Tác giả không chỉ giới thiệu các thư viện quan trọng như Pandas và NumPy mà còn hướng dẫn người đọc cách áp dụng chúng vào các bài toán thực tiễn. Những điểm nổi bật trong tài liệu bao gồm cách làm sạch dữ liệu, phân tích thống kê, và trực quan hóa dữ liệu, giúp người đọc nắm vững các kỹ năng cần thiết để trở thành một nhà phân tích dữ liệu thành công.

Để mở rộng kiến thức của bạn về lĩnh vực này, bạn có thể tham khảo tài liệu Dự án kết thúc học phần môn trực quan hoá và hệ thống thông tin địa lý đề tài biên dịch sách tài liệu python for data analysis, nơi bạn sẽ tìm thấy các ứng dụng thực tế của Python trong phân tích dữ liệu. Ngoài ra, tài liệu Nguyn van tun phan tich s liu va v cũng cung cấp hướng dẫn chi tiết về phân tích số liệu và tạo biểu đồ, giúp bạn có cái nhìn toàn diện hơn về các công cụ phân tích dữ liệu. Những tài liệu này sẽ là nguồn tài nguyên quý giá để bạn nâng cao kỹ năng và hiểu biết trong lĩnh vực phân tích dữ liệu.