AI Engineer (Full-Time or Contract)
We are looking for an experienced AI Engineer to work on real-world AI projects involving large language models, document intelligence, and private AI infrastructure.
Key Responsibilities:
- Deploy and optimize open-source Large Language Models (LLMs)
- Design, implement, and maintain private (on-premise) AI systems
- Develop Retrieval-Augmented Generation (RAG) architectures
- Build AI-powered pipelines for processing PDFs and other document formats
- Configure GPU-based inference environments and local AI infrastructure
- Develop AI-driven document classification and information extraction solutions
- Automate the analysis of legal and business documents using AI
- Build AI systems capable of generating response letters based on historical data and similar cases
Required Qualifications:
- Hands-on experience with AI and Large Language Model (LLM) technologies
- Strong proficiency in Python
- Practical experience with RAG, embeddings, and vector databases
- Experience in document classification and document understanding
- Knowledge of OCR technologies and document processing pipelines
- Strong prompt engineering skills and experience with AI agents
- Experience deploying and managing on-premise AI environments
- Experience with GPU inference, model deployment, and performance optimization
Preferred Qualifications:
- Experience with technologies such as vLLM, Ollama, Hugging Face Transformers, LangChain, LlamaIndex, or similar frameworks
- Experience working with NVIDIA GPUs and the CUDA ecosystem
- Experience in LegalTech or document automation projects
- Familiarity with Docker, Linux, and Microsoft Azure environments