- LLM Acceleration
- Efficient Inference
- Multimodal Large Models
- AI for Scientific
I am currently interested in practical acceleration methods for large language models and multimodal models, especially inference-side optimization, token-level compression, and efficient AI systems.
An offline-first scientific writing workspace, designed for researchers and students working with LaTeX, Python, and scientific documents.




