Description
Job Summary:
At Xpand IT, the Data Science team is dedicated to transforming data into impactful solutions through advanced modeling and industry-leading algorithms.
Key Highlights:
1. Participate in end-to-end Data Science projects
2. Lead R&D initiatives in GenAI and LLMs
3. Collaborate with engineering and business teams
**Team:** Data Science
**Level:** Mid\-level
**Offices:** Lisbon, Braga
At Xpand IT, the Data Science team is dedicated to transforming data into impactful solutions through advanced modeling and industry-leading algorithms. We tackle complex problems—from optimization and forecasting to recommendation systems—ensuring scalable, robust solutions deployed in real-world contexts. Our work goes beyond model creation: it includes building pipelines, integrating with existing systems, and collaborating closely with engineering and business teams.
Key Responsibilities
As a Data Scientist, you will participate in end-to-end projects—from raw data collection and processing to model deployment in production. You will also have a strong research component, responsible for exploring the market and emerging technologies. Your focus will include:
Developing and training predictive models, recommendation systems, and Machine Learning algorithms;
Leading R\&D initiatives by conducting research and developing PoCs to test new technologies and approaches (e.g., GenAI and LLMs);
Building and maintaining batch and near real\-time data and machine learning pipelines, and evaluating trade\-offs among performance, cost, and complexity;
**Ensure full project lifecycle management:** data extraction, preparation, manipulation, and model optimization;
Collaborating with engineering and business teams to translate real-world problems into data-driven solutions.
**Tech Stack:**
**Primary language:** Python (Pandas, NumPy, Scikit\-learn);
**Deep learning:** TensorFlow, Keras, PyTorch;
**Data \& ML:** PySpark, MLflow, Airflow, Azure ML;
**GenAI:** LLMs, Llama;
**Cloud:** Azure, Google Cloud or AWS;
**Exploration:** SQL and data visualization tools.
Requirements
Academic Qualifications
Bachelor’s or Master’s degree in Computer Engineering, Mathematics, Data Science, or related fields.
Data Science Experience
Solid experience (minimum 1 year) in developing and training Machine Learning models and Data Mining algorithms.
Proficiency in Python
Strong command of the Python ecosystem for data and mathematics (NumPy, SciPy, Pandas, Scikit\-learn).
End-to-End Project Experience
Practical experience across all phases of a data project—including raw data preparation and database manipulation (SQL).
Research \& Experimentation
Ability to conduct technical research, build proofs of concept (PoCs), and evaluate new tools or frameworks.
Languages
Fluency in both Portuguese and English (written and spoken) is mandatory.
**Bonus Points:**
Practical experience with GenAI, LLMs, and Prompt Engineering techniques;
Strong statistical knowledge (regressions, distributions, normality tests);
Experience with AI services in public cloud environments (Azure, Google Cloud or AWS);
Passion for sharing technical knowledge and staying up-to-date with market trends.