How to Manage Nested JSON Objects as a DataFrame in Pandas?

Front page > Programming > How to Manage Nested JSON Objects as a DataFrame in Pandas?

How to Manage Nested JSON Objects as a DataFrame in Pandas?

Published on 2024-11-08

Browse:845

How to Manage Nested JSON Objects as a DataFrame in Pandas?

Reading Nested JSON with Nested Objects as a Pandas DataFrame

When dealing with JSON data containing nested objects, manipulating it efficiently in Python is crucial. Pandas provides a powerful tool to achieve this - json_normalize.

Expanding the Array into Columns

To expand the locations array into separate columns, use json_normalize as follows:

import json
import pandas as pd

with open('myJson.json') as data_file:
    data = json.load(data_file)

df = pd.json_normalize(data, 'locations', ['date', 'number', 'name'], record_prefix='locations_')

print(df)

This will create a dataframe with expanded columns:

  locations_arrTime locations_arrTimeDiffMin locations_depTime  \
0                                                        06:32   
1             06:37                        1             06:40   
2             08:24                        1                     

  locations_depTimeDiffMin           locations_name locations_platform  \
0                        0  Spital am Pyhrn Bahnhof                  2   
1                        0  Windischgarsten Bahnhof                  2   
2                                    Linz/Donau Hbf               1A-B   

  locations_stationIdx locations_track number    name        date  
0                    0          R 3932         R 3932  01.10.2016  
1                    1                         R 3932  01.10.2016  
2                   22                         R 3932  01.10.2016

Handling Multiple JSON Objects

For JSON files containing multiple objects, the approach depends on the desired data structure.

Keep Individual Columns

To keep individual columns (date, number, name, locations), use the following:

df = pd.read_json('myJson.json')
df.locations = pd.DataFrame(df.locations.values.tolist())['name']
df = df.groupby(['date', 'name', 'number'])['locations'].apply(','.join).reset_index()

print(df)

This will group the data and concatenate the locations:

        date    name number                                          locations
0  2016-01-10  R 3932         Spital am Pyhrn Bahnhof,Windischgarsten Bahnho...

Flatten the Data Structure

If you prefer a flattened data structure, you can use json_normalize with the following settings:

df = pd.read_json('myJson.json', orient='records', convert_dates=['date'])

print(df)

This will output the data in a single table:

  number    date                   name  ... locations.arrTimeDiffMin locations.depTimeDiffMin locations.platform
0             R 3932  2016-01-10  R 3932  ...                       0                         0                  2
1             R 3932  2016-01-10  R 3932  ...                       1                         0                  2
2             R 3932  2016-01-10  R 3932  ...                       1                         -                  1A-B

Release Statement This article is reprinted at: 1729739643 If there is any infringement, please contact [email protected] to delete it

Latest tutorial More>

Is There a Performance Difference Between Using a For-Each Loop and an Iterator for Collection Traversal in Java?
For Each Loop vs. Iterator: Efficiency in Collection TraversalIntroductionWhen traversing a collection in Java, the choice arises between using a for-...

Programming Posted on 2025-03-11
How Can I UNION Database Tables with Different Numbers of Columns?
Combined tables with different columns] Can encounter challenges when trying to merge database tables with different columns. A straightforward way i...

Programming Posted on 2025-03-11
How to upload files with additional parameters using java.net.URLConnection and multipart/form-data encoding?
Uploading Files with HTTP RequestsTo upload files to an HTTP server while also submitting additional parameters, java.net.URLConnection and multipart/...

Programming Posted on 2025-03-11
$Why Isn\'t My CSS Background Image Appearing?$
Why Isn\'t My CSS Background Image Appearing?
Troubleshoot: CSS Background Image Not AppearingYou've encountered an issue where your background image fails to load despite following tutorial i...

Programming Posted on 2025-03-11
$Why Doesn\'t Firefox Display Images Using the CSS `content` Property?$
Why Doesn\'t Firefox Display Images Using the CSS `content` Property?
Displaying Images with Content URL in FirefoxAn issue has been encountered where certain browsers, specifically Firefox, fail to display images when r...

Programming Posted on 2025-03-11
How to Check if an Object Has a Specific Attribute in Python?
Method to Determine Object Attribute ExistenceThis inquiry seeks a method to verify the presence of a specific attribute within an object. Consider th...

Programming Posted on 2025-03-11
Why Does Microsoft Visual C++ Fail to Correctly Implement Two-Phase Template Instantiation?
The Mystery of "Broken" Two-Phase Template Instantiation in Microsoft Visual C Problem Statement:Users commonly express concerns that Micro...

Programming Posted on 2025-03-11
How do you extract a random element from an array in PHP?
Random Selection from an ArrayIn PHP, obtaining a random item from an array can be accomplished with ease. Consider the following array:$items = [523,...

Programming Posted on 2025-03-11
How to Write Truly Non-Blocking Functions in Node.js?
Correct Way to Write a Non-Blocking Function in Node.jsThe non-blocking paradigm is crucial in Node.js for achieving high performance. However, it can...

Programming Posted on 2025-03-10
How to Extract Text from Specific HTML Tags Using DOMDocument and XPath?
Parsing HTML with PHP's DOMDocument and XPathWhen attempting to parse HTML using PHP's DOMDocument, a common issue is finding specific text wi...

Programming Posted on 2025-03-10
d[IA]gnosis: developing RAG applications with IRIS for Healt
With the introduction of vector data types and the Vector Search functionality in IRIS, a whole world of possibilities opens up for the development of...

Programming Posted on 2025-03-10
Can We Create Generic Arrays in Java That Extend Comparable?
Generic Arrays in Java: Exploring Covariance and Type ErasureIntroductionGeneric arrays, where the array elements share a common type parameter, prese...

Programming Posted on 2025-03-09
$Why Does My WordPress Ajax Call Return \"0\"?$
Why Does My WordPress Ajax Call Return \"0\"?
Troubleshooting Ajax Calls in WordPress: Why Your Output is "0"In WordPress, making Ajax calls can be straightforward, but sometimes issues ...

Programming Posted on 2025-03-07
Can I Control the Height of Images Within CSS :before/:after Pseudo-Elements?
Can I Adjust Image Height in CSS :before/:after Pseudo-Elements?Your inquiry is whether it's possible to modify the height of an image used within...

Programming Posted on 2025-03-07
My Laravel Package Building Workflow
Crafting Laravel Packages: A Comprehensive Guide This article delves into the process of building Laravel packages, offering a structured approach fro...

Programming Posted on 2025-03-07