Overview
Analog is a catalyst for change across various industries - from global enterprises to the public sector, from healthcare to ecology. Analog's approach to AI ensures that technological advancements are not just about efficiency and convenience, but about enriching human experiences and fostering a more connected and mixed world. Analog is not just building technology; we are crafting a world where technology is an invisible yet indispensable part of our lives. Join us on our journey as we work to build solutions that will empower people to stop living behind the digital world and start remembering how to live in our rainbow-filled analog world.
Role
The Speech & Language Scientist is a specialized role that focuses on advancing the understanding and application of speech and language technologies. This role bridges research and development, enabling cutting-edge innovations in areas such as automatic speech recognition (ASR), natural language understanding (NLU), speech synthesis (TTS), and conversational AI.
Responsibilities
- Conduct cutting-edge research in speech and audio technologies, including areas such as speech recognition, speech synthesis, audio enhancement, speaker identification, and spatial audio
- Research, model, design, develop and test novel audio and speech processing algorithms using machine learning, signal processing, and computer vision
- Collaborate with cross-functional teams (e.g., software engineers, product managers, and designers) to implement research into real-world applications.
- Improve accuracy of models/algorithms to solve problems encountered during the model/algorithm implementation. Coordinate various teams to solve these problems in an efficient, and effective way.
- Propose scenarios where AI techniques could improve the efficiency of target clients’ businesses.
- Invent, analyze, design, develop and optimize algorithms and models in AUDIO areas such as Speech Recognition, Text to Speech, Speaker Recognition, etc.
Qualifications
- Ph.D. in Computer Science, Electrical Engineering, or a related field with a focus on speech and audio processing, or equivalent practical experience.
- Proven track record of research in speech or audio processing, demonstrated by publications in peer- reviewed journals or conferences.
- Expertise in machine learning frameworks (e.g., PyTorch, TensorFlow) and proficiency in programming languages like Python, C++, or similar.
- Strong knowledge of signal processing techniques and deep learning models for speech and audio tasks.
- Experience working with large-scale datasets and distributed computing systems.
- Proven track record of research in speech or audio processing, demonstrated by publications in peer- reviewed journals or conferences such as CVPR, ECCV, NIPS, ICASSP, InterSpeech etc
Preferred Qualifications
- Experience with real-time audio systems or applications in AR/VR, accessibility, or social interaction platforms.
- Familiarity with techniques for multilingual or low-resource speech recognition.
- Background in self-supervised or unsupervised learning for speech/audio representation.
- Knowledge of privacy-preserving techniques in audio data processing.
- Experience mentoring junior researchers or managing collaborative projects.
What Working At Analog Offers
- Culture: A collaborative, globally minded team building the future of physical intelligence. We value curiosity, courage, and creativity, and we’re united by a belief that technology should amplify human potential rather than replace it.
- Impact: The opportunity to work on projects with national and global significance, from adaptive cities and resilient infrastructure to next generation sports, entertainment, healthcare, and energy. Every product you help build contributes to shaping a safer, smarter, more human centered world.
- Growth: Outstanding opportunities for learning and career development. You’ll collaborate with leaders across AI, robotics, mixed reality, and systems engineering, while working on industry first, frontier scale solutions.
- Rewards: A competitive compensation package with healthcare, education support, and generous leave benefits, reflecting the high value we place on our people and their families.
نظرة عامة
تعد Analog محفزًا للتغيير عبر مختلف القطاعات - بدءًا من المؤسسات العالمية وحتى القطاع العام، ومن الرعاية الصحية إلى البيئة. يضمن نهج Analog في الذكاء الاصطناعي أن التقدم التكنولوجي لا يقتصر على الكفاءة والراحة فحسب، بل يتعلق بتمكين التجارب البشرية وإثراء العالم وتعزيز بيئة أكثر ترابطًا وتنوعًا. لا تقتصر Analog على بناء التكنولوجيا؛ بل نحن نصنع عالمًا تكون فيه التكنولوجيا جزءًا خفيًا ولكنه لا غنى عنه في حياتنا. انضم إلينا في رحلتنا بينما نعمل على بناء حلول تمكّن الناس من التوقف عن العيش خلف العالم الرقمي والبدء في تذكر كيفية العيش في عالمنا التناظري المليء بألوان الطيف.
الدور الوظيفي
يعد دور عالم الكلام واللغة دورًا متخصصًا يركز على تطوير الفهم والتطبيق لتقنيات الكلام واللغة. يربط هذا الدور بين البحث والتطوير، مما يتيح الابتكارات الرائدة في مجالات مثل التعرف الآلي على الكلام (ASR)، والفهم الطبيعي للغة (NLU)، وتوليد الكلام (TTS)، والذكاء الاصطناعي التفاعلي.
المسؤوليات
- إجراء أبحاث متطورة في تقنيات الكلام والصوت، بما في ذلك مجالات مثل التعرف على الكلام، وتوليد الكلام، وتحسين الصوت، والتعرف على المتحدث، والصوت المكاني
- بحث ونمذجة وتصميم وتطوير واختبار خوارزميات جديدة لمعالجة الصوت والكلام باستخدام التعلم الآلي ومعالجة الإشارات والرؤية الحاسوبية
- التعاون مع فرق متعددة التخصصات (مثل مهندسي البرمجيات ومديري المنتجات والمصممين) لتطبيق الأبحاث في تطبيقات العالم الحقيقي.
- تحسين دقة النماذج/الخوارزميات لحل المشكلات التي تواجهها أثناء تنفيذ النموذج/الخوارزمية. التنسيق بين الفرق المختلفة لحل هذه المشكلات بطريقة فعالة وكفؤة.
- اقترح سيناريوهات يمكن لتقنيات الذكاء الاصطناعي من خلالها تحسين كفاءة أعمال العملاء المستهدفين.
- ابتكار وتحليل وتصميم وتطوير وتحسين الخوارزميات والنماذج في مجالات الصوت مثل التعرف على الكلام، وتحويل النص إلى كلام، والتعرف على المتحدث، وما إلى ذلك.
المؤهلات
- درجة الدكتوراه في علوم الكمبيوتر، الهندسة الكهربائية، أو مجال ذي صلة مع التركيز على معالجة الكلام والصوت، أو خبرة عملية مكافئة.
- سجل حافل ومثبت من الأبحاث في معالجة الكلام أو الصوت، ويتجلى ذلك من خلال المنشورات في المجلات أو المؤتمرات المحكّمة.
- خبرة في أطر التعلم الآلي (مثل PyTorch وTensorFlow) وإتقان لغات البرمجة مثل Python وC++ أو ما يماثلها.
- معرفة قوية بتقنيات معالجة الإشارات ونماذج التعلم العميق لمهام الكلام والصوت.
- خبرة في العمل مع مجموعات البيانات ضخمة النطاق وأنظمة الحوسبة الموزعة.
- سجل حافل ومثبت من الأبحاث في معالجة الكلام أو الصوت، ويتجلى ذلك من خلال المنشورات في المجلات أو المؤتمرات المحكّمة مثل CVPR وECCV وNIPS وICASSP وInterSpeech وما إلى ذلك.
المؤهلات المفضلة
- خبرة في أنظمة أو تطبيقات الصوت بالوقت الفعلي في مجالات الواقع المعزز/الواقع الافتراضي (AR/VR)، أو إمكانية الوصول، أو منصات التفاعل الاجتماعي.
- إلمام بتقنيات التعرف على الكلام متعدد اللغات أو للغات شحيحة الموارد.
- خلفية في التعلم الذاتي أو غير الخاضع للإشراف لتمثيل الكلام/الصوت.
- معرفة بالتقنيات الحافظة للخصوصية في معالجة البيانات الصوتية.
- خبرة في توجيه الباحثين المبتدئين أو إدارة المشاريع التعاونية.
ما تقدمه بيئة العمل في Analog
- الثقافة: فريق تعاوني بعقلية عالمية يبني مستقبل الذكاء المادي. نحن نثمن الفضول والشجاعة والإبداع، ويجمعنا الإيمان بأنه ينبغي للتكنولوجيا أن تعزز الإمكانات البشرية لا أن تحل محلها.
- الأثر: فرصة العمل في مشاريع ذات أهمية الوطنية وعالمية، بدءًا من المدن التكيفية والبنية التحتية المرنة إلى الجيل القادم من الرياضة والترفيه والرعاية الصحية والطاقة. كل منتج تساعد في بنائه يساهم في تشكيل عالم أكثر أمانًا وذكاءً وأكثر تركيزًا على الإنسان.
- النمو: فرص استثنائية للتعلم والتطوير المهني. ستتعاون مع قادة عبر الذكاء الاصطناعي، الروبوتات، الواقع المختلط، وهندسة الأنظمة، أثناء العمل على حلول رائدة على مستوى الحدود التقنية.
- المكافآت: حزمة تعويضات تنافسية تتضمن الرعاية الصحية، ودعم التعليم، ومزايا إجازات سخية، مما يعكس القيمة العالية التي نوليها لموظفينا وعائلاتهم.