Hacktoberfest 2026: the issues maintainers tagged for October, open and beginner-friendly. Browse Hacktoberfest issues

register_pandas_dataframe crashes kernel when receiving dataframe with multiple columns with same name

Open
#1,809 0 comments 2 reactions 0 assignees View on GitHub

Nobody has claimed this yet.

Assessment

Difficulty
4/5
Estimated time
3-5 days
Newbie friendliness
30/100
Issue type
Bug
Clarity
Mostly clear
Activity status
Stale
Tech stack
azure, jupyter-notebook, pandas, python

Research direction

Start by running the pandas DataFrame example from the issue with duplicate column names and observe the kernel exit. The payload identifies the TabularDatasetFactory documentation source at AzureML-Docset/stable/docs-ref-autogen/azureml-core/azureml.data.dataset_factory.tabulardatasetfactory.yml, but no implementation file or test is named. Done means the failure returns useful feedback without crashing the IPython kernel.

Written by the indexing model from the issue text.

Description

My Synapse kernel crashed when I tried to register a dataset where the dataframe had multiple columns with the same name. I am not completely sure if this is the correct place to post this, but this is the best place I could find.

Minimum reproducible example:

dataset = Dataset.Tabular.register_pandas_dataframe(
    dataframe = pd.DataFrame([[1,2],[1,2]], columns=['a', 'a']),
    target = datastore,
    name = "bad_dataframe",
    show_progress=True,
)

What I would expect:
To get some feedback about why it fails, and the kernel should not crash.

What I get:

Validating arguments.
Arguments validated.
Successfully obtained datastore reference and path.
Uploading file to managed-dataset/13e4b3c6-8730-4951-bde5-2d2bc6456e2d/
InternalError: Ipython kernel exits with code -11. Please restart your session.

Document Details

Do not edit this section. It is required for docs.microsoft.com ➟ GitHub issue linking.

Dominant language
Jupyter Notebook
Stars
4.4k
Forks
2.6k
PR merge metrics
No merged PRs in 30d

Contributor guide

No contributing guide indexed for this repository

First steps

  1. Read the whole issue, then the project's contributing guide.
  2. Comment on the issue to say you are picking it up — it saves two people doing the same work.
  3. Fork the repository and make your change on a branch.
  4. Open a pull request that references the issue number.

More from Azure/MachineLearningNotebooks

All issues in Azure/MachineLearningNotebooks

Similar issues

More Data Engineering issues

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.