BEGIN:VCALENDAR
VERSION:2.0
PRODID:-//ENCCS - ECPv6.17.2//NONSGML v1.0//EN
CALSCALE:GREGORIAN
METHOD:PUBLISH
X-WR-CALNAME:ENCCS
X-ORIGINAL-URL:https://enccs.se
X-WR-CALDESC:Events for ENCCS
REFRESH-INTERVAL;VALUE=DURATION:PT1H
X-Robots-Tag:noindex
X-PUBLISHED-TTL:PT1H
BEGIN:VTIMEZONE
TZID:Europe/Stockholm
BEGIN:DAYLIGHT
TZOFFSETFROM:+0100
TZOFFSETTO:+0200
TZNAME:CEST
DTSTART:20250330T010000
END:DAYLIGHT
BEGIN:STANDARD
TZOFFSETFROM:+0200
TZOFFSETTO:+0100
TZNAME:CET
DTSTART:20251026T010000
END:STANDARD
BEGIN:DAYLIGHT
TZOFFSETFROM:+0100
TZOFFSETTO:+0200
TZNAME:CEST
DTSTART:20260329T010000
END:DAYLIGHT
BEGIN:STANDARD
TZOFFSETFROM:+0200
TZOFFSETTO:+0100
TZNAME:CET
DTSTART:20261025T010000
END:STANDARD
BEGIN:DAYLIGHT
TZOFFSETFROM:+0100
TZOFFSETTO:+0200
TZNAME:CEST
DTSTART:20270328T010000
END:DAYLIGHT
BEGIN:STANDARD
TZOFFSETFROM:+0200
TZOFFSETTO:+0100
TZNAME:CET
DTSTART:20271031T010000
END:STANDARD
END:VTIMEZONE
BEGIN:VEVENT
DTSTART;TZID=Europe/Stockholm:20260915T090000
DTEND;TZID=Europe/Stockholm:20260916T120000
DTSTAMP:20260702T142822Z
CREATED:20260702T142822Z
LAST-MODIFIED:20260702T142822Z
UID:38708-1789462800-1789560000@enccs.se
SUMMARY:[Workshop] Practical Data Wrangling
DESCRIPTION:Overview\nData is essential in data-driven projects\, as it forms the foundation for all subsequent analysis\, modeling\, and decision-making. Depending on the specific task\, raw data may be collected from a wide variety of sources such as databases\, APIs\, sensors\, logs\, documents\, or images. Before it can be effectively used for analysis or machine learning\, raw data must be cleaned\, transformed\, validated\, and organized into a consistent and usable format. As data comes in many different forms\, including numerical\, categorical\, time series\, text\, event/log\, and image data\, the tools and techniques used for data wrangling can vary significantly depending on the data type and the requirements of the task. \nIn this workshop\, we will cover practical data wrangling techniques for numerical\, categorical\, time series\, text\, event/log\, and image data. Participants will learn how to detect and handle missing values\, outliers\, inconsistencies\, duplicates\, and formatting problems. We will demonstrate methods for transforming and encoding categorical variables\, parsing and aggregating temporal data\, processing unstructured text\, analyzing event logs\, and preparing image datasets for machine learning and analytics. Each session combines concepts\, demonstrations\, and hands-on exercises using realistic datasets to help participants develop practical skills that can be applied immediately in downstream modeling tasks. \nWho is this workshop for?\nThis workshop is designed for \n\ndata practitioners who regularly work with raw or semi-structured data and need to prepare it for analysis or modeling.\ndata analysts\, data scientists\, machine learning engineers\, and software engineers who want to strengthen their practical data preprocessing skills.\ngraduate students and researchers working with real-world datasets who need a structured approach to data cleaning and transformation.\n\nPrerequisites\nTo ensure a smooth learning experience\, participants should have: \n\nbasic proficiency in Python programming (variables\, loops\, functions) and some libraries like NumPy\, Pandas\, and Matplotlib/Seaborn.\nbasic familiarity with statistics (mean\, median\, variance) and introductory machine learning concepts will make it easier to follow the examples.\nbe comfortable reading and writing simple code and working with datasets in a notebook environment.\n\nKey Takeaways\nIn this workshop\, participants will learn how to systematically clean and structure different types of real-world data\, including tabular\, time series\, text\, event/log\, and image data.\nBy the end of this workshop\, participants will: \n\ngain practical experience in building reproducible data wrangling pipelines that improve data quality and usability for downstream data analysis and machine learning.\nunderstand common pitfalls in messy datasets and how to address them effectively using standard techniques and tools.\nbe able to confidently transform raw datasets into well-structured inputs suitable for downstream modeling and analysis tasks.\n\nTentative Schedule (TBA)\nDay 1 \n\nIntroduction\nData Types and Data Storage Formats\nNumerical Data Wrangling\nCategorical Data Wrangling\nTime Series Data Wrangling\n\nDay 2 \n\nText Data Wrangling\nEvent and Log Data\, and Image Data Wrangling\nWrangling Other Data Types\nSummary and Key Takeaways\n\nRegulations\nDue to EuroCC3 regulations\, we CAN NOT ACCEPT generic or private email addresses. Please use your official university or company email address for registration. \nThis training is for users that live and work in the European Union or a country associated with Horizon 2020. You can read more about the countries associated with Horizon2020 HERE. \nContact\nFor questions regarding this workshop or general questions about ENCCS training events\, please contact training@enccs.se.
URL:https://enccs.se/event/workshop-practical-data-wrangling/
LOCATION:Online
ATTACH;FMTTYPE=image/jpeg:https://media.enccs.se/2026/07/practical-data-wrangling.jpg
END:VEVENT
END:VCALENDAR