تعد IC Markets Global واحدة من أشهر شركات تقديم خدمات تداول عقود الفروقات على الفوركس، حيث توفر حلول تداول للمتداولين اليوميين النشطين والمتداولين بأسلوب المضاربة السريعة (سكالبينج)، بالإضافة إلى المتداولين الجدد في سوق الفوركس. وتقدم IC Markets Global لعملائها منصات تداول متطورة، واتصالاً منخفض زمن الاستجابة، وسيولة ممتازة.
حدثت IC Markets Global ثورة في عالم تداول الفوركس عبر الإنترنت. بات بإمكان المتداولين الآن الحصول على أسعار كانت تقتصر في السابق على البنوك الاستثمارية وأصحاب الثروات الكبيرة.
يمتلك فريق الإدارة لدينا خبرة كبيرة في أسواق الفوركس وعقود الفروقات والأسهم في آسيا وأوروبا وأمريكا الشمالية. وهذه الخبرة هي التي مكنتنا من اختيار أفضل الحلول التكنولوجية الممكنة وانتقاء أفضل مزودي الأسعار المتاحين في السوق.
عن الوظيفة
نبحث عن مهندس أول عمليات نماذج اللغات الكبيرة (LLMOps) / منصة الذكاء الاصطناعي لبناء ونشر وتحسين وتشغيل البنية التحتية لنماذج اللغات الكبيرة والذكاء الاصطناعي التوليدي الجاهزة للإنتاج. يجمع هذا الدور بين استدلال LLM، وتحسين وحدات معالجة الرسومات (GPU)، وKubernetes، والبنية التحتية السحابية، القابلية للملاحظة والمراقبة (Observability)، والتوليد المعزز بالاسترجاع (RAG)، وهندسة منصات الذكاء الاصطناعي.
المسؤوليات الرئيسية
- نشر وتشغيل نماذج اللغات الكبيرة (LLMs) المستضافة ذاتياً باستخدام vLLM وSGLang وOllama.
- تحسين استدلال LLM من حيث زمن الاستجابة، وإنتاجية البيانات، والتزامن، وذاكرة GPU، والتخزين المؤقت KV، والتكلفة.
- إدارة أعباء عمل GPU عبر العديد من وحدات معالجة الرسومات من NVIDIA.
- نشر وصيانة خدمات الذكاء الاصطناعي على Kubernetes / AWS EKS باستخدام Docker وHelm.
- تطبيق آليات موثوقية LLM بما في ذلك فحوصات السلامة، والمراقبة، والتعافي الآلي، واستراتيجيات إعادة تشغيل/تحديث النماذج.
- تطبيق القابلية للملاحظة والمراقبة باستخدام Langfuse/LangSmith وOpenTelemetry وPrometheus وGrafana.
- نشر وتحسين أنظمة RAG، ونماذج التضمين (Embedding)، وقواعد البيانات المتجهة مثل Qdrant وMilvus.
- دعم وكلاء الذكاء الاصطناعي وسير العمل المبني باستخدام LangChain وLangGraph.
- بناء وصيانة أنابيب CI/CD لخدمات الذكاء الاصطناعي والبنية التحتية.
- استكشاف أخطاء بيئة الإنتاج وإصلاحها عبر نماذج LLM ووحدات GPU وKubernetes والشبكات وتطبيقات الذكاء الاصطناعي.
المهارات المطلوبة
- خبرة قوية في Python وFastAPI
- vLLM وHugging Face ونشر LLM المستضاف ذاتياً
- Kubernetes وDocker وHelm وAWS
- استدلال NVIDIA GPU وتحسين الأداء
- LangChain / LangGraph / LangSmith
- RAG والتضمينات وقواعد البيانات المتجهة (Qdrant)
- القابلية للملاحظة ومراقبة نماذج LLM
- PostgreSQL / Redis
- GitHub Actions / CI/CD
- مهارات قوية في استكشاف أخطاء بيئة الإنتاج وإصلاحها
IC Markets Global is one of the most renowned Forex CFD provider, offering trading solutions for active day traders and scalpers as well as traders that are new to the forex market. IC Markets Global offers its clients cutting edge trading platforms, low latency connectivity and superior liquidity.
IC Markets Global is revolutionizing online forex trading. Traders are now able to gain access to pricing previously only available to investment banks and high net worth individuals.
Our management team have significant experience in the Forex, CFD and Equity markets in Asia, Europe and North America. It is this experience that has enabled us to select the best possible technology solutions and hand pick some of the best pricing providers available in the market.
About the Role
We are looking for a Senior LLMOps / AI Platform Engineer to build, deploy, optimize, and operate production-grade LLM and Generative AI infrastructure. The role combines LLM inference, GPU optimization, Kubernetes, cloud infrastructure, observability, RAG, and AI platform engineering.
Key Responsibilities
- Deploy and operate self-hosted LLMs using vLLM, SGLang, Ollama.
- Optimize LLM inference for latency, throughput, concurrency, GPU memory, KV cache, and cost.
- Manage GPU workloads across multiple NVIDIA GPUs.
- Deploy and maintain AI services on Kubernetes / AWS EKS using Docker and Helm.
- Implement LLM reliability mechanisms including health checks, monitoring, automated recovery, and model restart/refresh strategies.
- Implement observability using Langfuse/LangSmith, OpenTelemetry, Prometheus, and Grafana.
- Deploy and optimize RAG systems, embedding models, and vector databases such as Qdrant, Milvus.
- Support AI agents and workflows built with LangChain and LangGraph.
- Build and maintain CI/CD pipelines for AI services and infrastructure.
- Troubleshoot production issues across LLMs, GPUs, Kubernetes, networking, and AI applications.
Required Skills
- Strong Python and FastAPI experience
- vLLM, Hugging Face and self-hosted LLM deployment
- Kubernetes, Docker, Helm and AWS
- NVIDIA GPU inference and performance optimization
- LangChain / LangGraph/LangSmith
- RAG, embeddings and vector databases (Qdrant)
- LLM observability and monitoring
- PostgreSQL / Redis
- GitHub Actions / CI/CD
- Strong production troubleshooting skills