Jerome Nilmeier

Data Scientist & Developer Advocate, IBM

Jerome Nilmeier is a developer advocate, data scientist, and member of the IBM Center for Open source Data and AI Technologies (CODAIT), where he works with with open source frameworks for big data, machine learning, and deep learning as a developer advocate. He has recently published an O'Reilly Manual, "Data Science and Engineering at Enterprise Scale", which is a great introductory text for data scientists interested in machine learning, big data, and AI. He has a BS in Chemical Engineering from UC Berkeley, a PhD in Computational Biophysics from UC San Francisco, and has carried out postdoctoral research in biophysics and bioinformatics at UC Berkeley, Lawrence Berkeley and Livermore Laboratories, and at Stanford as an OpenMM Fellow. Just prior to joining IBM, he completed the Insight Data Engineering program in late 2014. He has been with IBM since 2015.

Jerome Nilmeier

Sessions / 2019 / 2 talks

  • RNNs and LSTMs have enjoyed great success in text generation algorithms, but their use in other fields has not been as widely studied.  We will discuss our experiences and progress using Recurrent Neural Networks to make predictions on arbitrary multivariate time series data.  Our first study used weather data from the JFK terminal over several years using the TensorFlow framework.  We will discuss the issues related to tuning and validating this model, as well as how we migrated this model into the Model Asset Exchange, which is an IBM hosted API for making predictions on data using pre-trained neural network models.  Our insight into tuning this model allowed us to provide another API via Watson Machine Learning, which is a hosted service that allows user defined data and models to be uploaded, trained, and tuned on GPU accelerated on demand hardware using simple remote API calls.  We will discuss examples from the financial sector, weather prediction, and other important time series prediction use cases.

  • TensorFlow has emerged as one of the most popular deep learning frameworks in use today. It has by far more users and contributors than any other project, and appears to be continuing on its upward trajectory with the release of the TensorFlow 2.0 API. There are many new features of the TensorFlow 2.0 API to look through, and we will discuss many of them, including eager execution, notebook accessible tensorboards, and tighter integration with Keras. In addition, we will show how edge calculations can be accelerated with tensorflow.js which runs completely in the browser and provides much faster model serving. TensorFlow Lite is a framework for running on smaller remote devices. We will also discuss the parallel execution framework, as TensorFlow is well on the way to becoming the standard for all things deep learning.

Sessions / 2016 / 1 talk

  • Apache systemML is IBM's open source project that interfaces with the Spark Context, allowing for simple expression of numerical algorithms. This is an ideal platform for Data Science, especially when there is an interest in specializing machine learning algorithms for specific challenges. The platform is extremely flexible, and enables complex numerical algorithms to be expressed in a simple and readable syntax, while preserving scalability for heavy duty computations. The parallelization details are optimized through the powerful cost based optimization engine in systemML.

Ready to take this stage?

The next edition is programmed by practitioners. Tell us what you built and what you learned.

Apply to be a speaker