Our Research Projects

Reasoning-Aware Mixture-of-Experts for Large Language Models

This is a project which is currently making use of HPC facilities at Newcastle University. It is active.

Project Contacts

For further information about this project, please contact:


Project Description

This project investigates reasoning-aware Mixture-of-Experts (MoE) architectures for Large Language Models (LLMs). The research focuses on improving reasoning performance by studying expert specialisation, routing strategies, and architectural modifications for complex reasoning tasks. Experiments involve training and evaluating transformer-based MoE models on public reasoning benchmarks, including mathematical, commonsense, and multi-step reasoning datasets. The project also analyses expert routing behaviour and model efficiency to better understand the relationship between routing decisions and reasoning capability.