About
My focus narrows to transformer architectures and their fine-tuning for low-resource NLP. Ignore the general LLM hype; optimizing inference pipelines for specific, constrained environments is where the real work lies. It's all about model compression and efficient deployment.
No questions yet.