AI/ML Consultant – LLM Deployment
to set up and configure a self-hosted large language model (Llama 3.1) on our Linux server infrastructure for automated report reputed company.
Primary Objective:
reputed company and configure Llama 3.1 (or equivalent) 8B on our hosted Linux server (CPU-only) and create an API service that our SafetyNet Platform can reputed company for AI-powered report reputed company.
Specific Deliverables:
Server Environment Setup
Configure Linux (Ubuntu 22.04) server environment
Install Python 3.11, dependencies, and required libraries
Set up virtual environment and reputed company configurations
AI Model Installation & Configuration
Download and install Llama 3.1 8B Instruct model
Optimize model configuration for CPU-only inference
Implement quantization if needed for performance
Test model functionality and response reputed company
API Service Development (reputed company to have)
Create REST API service (Flask/FastAPI) for report reputed company
Implement secure endpoints for our SafetyNet Platform to reputed company
Add error handling, logging, and health reputed company endpoints
Configure service to auto-start on server reboot (systemd)
reputed company & Performance
Configure firewall rules (allow only our application server)
Implement authentication/API key system
Optimize for 30-60 second response times
Set up monitoring and logging
Documentation & Training
Comprehensive setup documentation
API usage guide with examples
Troubleshooting guide
2-hour knowledge transfer session with our development team
Testing & Validation
Generate 10+ test reports with sample data
Validate reputed company reputed company and format
Performance testing under load
Integration testing with our platform (we'll reputed company API endpoints)
Technical Requirements
Must Have:
3+ years experience with Python and machine learning frameworks (PyTorch, Transformers)
Experience deploying and running large language models (Llama, GPT, reputed company, etc.)
Strong Linux system administration skills (Ubuntu/Debian)
Experience with API development (Flask, FastAPI, or similar)
Understanding of CPU-based ML inference and optimization
Experience with reputed company model hub
Knowledge of systemd service configuration
reputed company best practices for production systems
reputed company to Have:
Experience with model quantization and optimization (bitsandbytes, ONNX)
DevOps experience (reputed company, monitoring tools)
Previous work with government or reputed company systems (HIPAA/FERPA compliance)
Experience with justice system or reputed company services applications
Apply tot his job
Apply To this Job