Location:
Toronto, Ontario (Initially Remote) Duties and Responsibilities:
Building the backend engine that runs our product. This includes extending our existing Machine Learning and Big Data pipelines and building entirely new capabilities, including:
Big Data cluster, workflows and applications: data pipelines at scale, and real-time processing Machine learning and Data Scientist support: used in linguistics, ranking, classification, and other artificial intelligence applications
Ingestion Pipeline: process data that comes from our web crawler which discovers and fetches content from the web and other sources Skills and Qualifications:
BS or MS degree in Computer Science Solid experience with Java programming (we also use Spring, Spring Webflux, Reactor, Netty) Multi-threading experience is a must Experience in scalable architectures and high-throughput application design Experience working with microservices/REST APIs (HTTP, XML, JSON) Comfortable in Linux and Windows environments. Prefer some experience with Big Data Technologies (at least one of the following):
Hadoop ecosystem (HDFS, Hadoop, Hive) Spark Samza Kafka Aerospike Lucene NLP (Solr or ElasticSearch)
Preferred Experience:
Continuous Integration and Deployment Tool usage: Git/Gitlab and IntelliJ Functional programming experience (Scala) Gradle The ideal candidate will be self‑motivated, possess excellent communication skills (both oral and written) and be able to work efficiently and independently in an agile development environment. We offer a full comprehensive benefits package including medical, dental and vision. Employees receive a generous time off (PTO) plan and 13 holidays per year. We also offer 401(k) benefits, long term disability benefits and life insurance. With a casual and flexible work environment.
#J-18808-Ljbffr