BEGIN:VCALENDAR
VERSION:2.0
PRODID:icalendar-ruby
CALSCALE:GREGORIAN
BEGIN:VEVENT
DTSTAMP:20260720T131752Z
UID:10cdd992-894b-4d47-a84e-f133ee0b21ca
DTSTART:20191031T100000Z
DTEND:20191101T173000Z
DESCRIPTION:Apache Spark is an open-source framework for cluster computing\
 , ideal for large-scale parallel data processing\, that is designed for pe
 rformance and ease-of-use. It is faster and simpler to use than Hadoop Map
 Reduce\, providing a rich set of APIs in Python\, Java and Scala.\n\nThis 
 hands-on course will cover the following topics:\n\n\n	Introduction to Spa
 rk\n	Map\, Filter and Reduce\n	Running on a Spark Cluster\n	Key-value pair
 s\n	Correlations\, logistic regression\n	Decision trees\, K-means\n\n\nSes
 sions\n\n10:00 - 17:30 (Thu)\n10:00 - 15:30 (Fri)\n\nAttendees will be pro
 vided with access to EPCC's Tier2 Cirrus system for all practical exercise
 s.\n\nThe practicals will be done using Jupyter notebooks so a basic knowl
 edge of Python would be extremely useful.\n\nRegistration: Registration ha
 s been closed as the course is full with a long waiting list.\n\nTimetable
 \n\nFull timetable and course materials. \n\n \nhttps://events.prace-ri.
 eu/event/922/
SUMMARY:Introduction to Spark for Data Scientists @ EPCC at Alan Turing Ins
 titute London
URL;VALUE=URI:https://events.prace-ri.eu/event/922/
END:VEVENT
END:VCALENDAR
