Job Search

Recruit Detail

Find out more about the work we do, the experience and skills we can bring to the table, and our terms and conditions.

Company Name

ELYZA Co., Ltd.

Job Type

1A09 [Research and Development] Senior Research Scientist (Voice-Based Model Development)

Work Detail

[Expected Role and Responsibilities] As a Senior Research Scientist, you will play a central role in conceiving the direction of ELYZA's research and development based on the company's mission, business strategy, and market environment, and consistently driving the process from research theme planning to experimental design, verification, and dissemination. Research Strategy Formulation: Planning and promoting research themes in the speech-based modeling domain, such as speech dialogue, speech recognition, and speech synthesis. Research and Development Promotion: Exploring and verifying learning methods for speech-based models that integrate linguistic knowledge and speech processing capabilities. Business Collaboration: Setting research topics and clarifying issues based on collaboration with other business units, with a view to social implementation and business impact. Team Leadership: Planning, promoting, and reviewing experimental plans involving research engineers and other professionals. Enhancing Presence: Enhancing ELYZA's research presence through diverse forms of external communication, such as papers, technical blogs, model releases, and presentations.

Ideal Profile

Required Skills [All of the following are required] Specialized knowledge in the fields of machine learning and deep learning, and 5+ years of research and development experience Practical experience in model implementation, experimentation, and evaluation using Python Communication skills to lead projects while collaborating with team members Business-level communication skills in Japanese / Native-level Japanese proficiency Preferred Skills A PhD in information science or a similar field, or equivalent experience Research and development experience in speech-based modeling such as speech recognition, speech synthesis, and speech dialogue Research and development experience in speech signal processing and speech feature extraction Acceptance of proposals at top conferences in computer science, machine learning, and speech technology Experience in research promotion and project management in industry-academia collaboration, joint research, and national projects Desired Candidate Profile ELYZA's core value: Long Term We are looking for individuals who share Greedy's philosophy, who can perform their duties with ownership and a strong sense of ethics, who can break down abstract issues, formulate hypotheses, and explore the optimal approach within limited resources, who can choose research topics with an awareness of their impact on the business, organization, and society, while discussing with engineers and business-side members, without being confined to team boundaries, who are motivated to involve others and achieve significant results as a team, and who are eager to develop and promote research strategies in new technological areas such as voice infrastructure models, while incorporating LLM knowledge.

What we especially look for in a PhD or postdoctoral fellow

A PhD in information science or a related field, or equivalent experience.

Work Location

[Head Office Location] ◆SWT Building 5th & 6th Floor, 3-15-9 Hongo, Bunkyo-ku, Tokyo [Work Location] ◆Head office or home (no change allowed) *No relocation required *Full remote work is possible (residence limited to within Japan)

Phd. Stating Salary

◆Expected Annual Salary: ¥9,000,000 - ¥15,000,000 ◆Monthly Salary: ¥625,000 - *Determined based on experience and abilities <Breakdown> Monthly Base Salary: ¥376,913 - Fixed Overtime Allowance (38 hours/month): ¥139,871 - Late-Night Overtime Allowance (19 hours/month): ¥13,988 - Skill Allowance: ¥94,228 - *Overtime pay for hours exceeding the fixed amount will be calculated in 1-minute increments ◆Performance Evaluation: Twice a year

Selection Flow

Document screening → Casual interview/HR interview → First interview → Second interview → Final interview → Offer/Proposal interview * After confirming your interest in the selection process, we will conduct 1-3 interviews before the final interview. * During the selection process, we will conduct casual interviews and, if requested, trial work experience to facilitate mutual understanding. * You can choose between online and offline interview formats. Office tours are also possible. * We will conduct a reference check. Details will be provided once you proceed to the selection stage.

Similar Recruits

CyberAgent, Inc.

Job Type
[AI Lab] (Research Engineer) We are looking for a research engineer in the field of voice dialogue!

Working in collaboration with research scientists and business units, you will leverage the knowledge accumulated through field testing and R&D in real-world environments such as commercial facilities and accommodations to implement an autonomous voice dialogue system that understands human behavior and speech and engages in appropriate voice interaction. You will then verify its effectiveness in real-world settings. In the future, as a technical specialist, you will develop a foundational system that can be used across a wide range of fields, while understanding the organization's challenges. ▼Main Tasks Development of a voice dialogue system using a dialogue architecture currently under research and development at the AI ​​Lab Implementation of recognition functions (image processing and voice processing) based on robot-human interaction Implementation of machine learning models on edge devices Construction of a system for deploying robot recognition function components to the edge via the cloud Discussion with researchers, proposal, implementation, and introduction of technologies and models Construction of a system for collecting analytical data for research Collaboration with industry-academia partners and joint research institutions, incorporating academic knowledge, and applying it to product development, problem-solving, and research paper analysis.

Fairy Devices Co., Ltd.

Job Type
Researcher (Speech and Language Processing)

In this position, you will be responsible for the research and development and functional enhancement of mimi, the foundational technology supporting our company's activities, as well as the research portion of collaborative research and development projects related to mimi based on client requests. mimi emphasizes being a foundational technology for realizing superior user experiences, and therefore integrates various core technologies necessary for this purpose. For example, the following research topics are anticipated: - Speech signal processing and speech language processing technologies such as speech recognition, speaker recognition, emotion recognition, speech enhancement, and sound source separation. - Natural language processing technologies for dialogue systems such as intent understanding, dialogue control, and response generation. In line with our goal of exploring new possibilities for voice solutions and voice dialogue systems, we will tackle a wide range of research topics, taking into account the progress of academic research and technological development, and are not necessarily limited to the research topics listed above. You will be required to achieve our goals through rational means, with a broad perspective that includes not only creating new technologies yourself, but also introducing technologies from other research institutions. At the same time, it is essential that all research and development personnel work together collaboratively, keeping in mind that we are creating a single core technology. Actively presenting research and development results at domestic and international academic conferences and in academic journals, as well as interacting with researchers outside the company, are also part of the job responsibilities.

Honda Research Institute Japan Co., Ltd.

Job Type
Researcher specializing in speech recognition and acoustic processing algorithms. * Fixed-term contract, renewable.

Established in 2003 in Japan, the United States, and Europe, as a wholly owned subsidiary of Honda R&D Co., Ltd., HRI aims to be an organization that challenges new fields beyond automotive technology. Our guiding principle is "Innovate through Science." Towards the realization of a hybrid society where people, the global environment, and intelligent systems coexist in harmony, we focus on communication and sensing, such as Cooperative Intelligence and Cooperative Devices, through a multifaceted approach drawing from artificial intelligence, robotics, systems science, neuroscience, materials science, psychology, and social ethics. We are engaged in various research projects. ■ Job Openings We are seeking a "Speech Recognition and Acoustic Processing Algorithm Researcher" to work on next-generation technology research to solve various problems related to acoustics and speech recognition. [Job Details] * Acoustic and language model adaptation in speech recognition, end-to-end models, dialization, and Kaldi/ESPnet applications * Time-series signal processing, speech enhancement, and noise reduction in acoustic processing * Strategic planning for acoustic and speech recognition research in collaboration with Honda Group companies and sister companies In the future, we expect you to become a team leader, responsible for formulating strategies in this field, proposing new technologies and research themes, project management, mentoring team members and junior colleagues, building networks with internal and external experts, and making external presentations. [Job Characteristics] This job involves working with cutting-edge technologies in next-generation acoustic and speech recognition research, from planning and proposing solutions to implementation and experimentation, with a view to applications. You will actively conduct research and make recommendations, including publishing your own papers. Leveraging your experience and knowledge, you will be able to plan and propose research directions, goals, approaches, and plans, engaging in challenging and rewarding work. • Speech Recognition Algorithms (Acoustic Models, Language Models, End-to-End Models, Dialysis, etc.) • Acoustic Processing Algorithms (Time-Series Signal Processing, Speech Enhancement, Noise Reduction, etc.) • Machine Learning Algorithms Related to the Above (Deep Learning, etc.) Based on Honda's shared philosophy that "technology exists for the benefit of people," we will leverage our experience and knowledge to lead research and development with speed and innovative ideas, aiming to become a unique entity in the world. We will grasp the cutting edge trends in the acoustics and speech recognition fields, clearly define the value of the technological research we should pursue, and proactively promote research on next-generation acoustics and speech recognition algorithms. Furthermore, we will envision specific application fields and users, and aim for higher results through rapid prototyping, experimentation, analysis, and feedback. We will actively disseminate the research results obtained externally through international conferences and papers, contributing to enhancing our company's presence. Promoted as a research project under the Research Division Manager (Department: Research Division). Research projects are initiated after the individual proposes the content and budget, and are approved by the board. • A highly international workplace with approximately half of all employees being foreign nationals. Communication within the company is conducted daily in both Japanese and English. • The workplace is located within Honda Motor Co., Ltd.'s "Wako Campus," a suburban office that has received numerous awards. • Researchers have considerable autonomy, and the external publication of research results is actively encouraged. There is a culture that respects the company's direction while also valuing the will of the researchers.

2026年卒