{"schemaVersion":"jobsearcher.job.v1","id":"95495c24b80cc83cc0d3e30a","url":"https://jobsearcher.com/jobs/95495c24b80cc83cc0d3e30a","canonicalUrl":"https://jobsearcher.com/jobs/95495c24b80cc83cc0d3e30a","title":"Visual Generation & Multimodal Evaluation Machine Learning Engineer Intern (AML-Ark-US) - 2027 [...]","description":"Join us as we work together to inspire creativity and enrich life around the globe.\nLocation\nSan Jose\nTeam\nTechnology\nEmployment Type\nIntern\nJob Code\nA174683\nShare this listing\nResponsibilities\nThe Applied Machine Learning Ark team combines system engineering and machine learning to develop and operate Large Language Model (LLM) service platforms that offer businesses Model-as-a-Service (MaaS) solutions, serving both large model providers and downstream users. The US team drives the design, development, and operation of MaaS solutions across the US and international markets outside mainland China. We are building full-stack, end-to-end solutions spanning text and multimodal LLM algorithms, LLM training/fine-tuning/inference frameworks, prompt engineering, model alignment, and intelligent agent systems. Beyond model serving, we operate large-scale log analytics pipelines that process massive volumes of invocation logs from text models, multimodal models, and agent systems — extracting usage patterns, quality signals, and actionable insights to inform model improvement, system optimization, and product decisions through continuous, data-driven feedback loops. We are actively seeking talented engineers and researchers specializing in Large Language Models and AI Agent systems to join our dynamic team. We are looking for talented individuals to join us for an internship. Our internship program offers students hands‑on experience, industry exposure, and opportunities to apply their knowledge to real‑world challenges while building a strong foundation for personal and professional growth. Interns will gain practical experience, explore potential career paths, and participate in social events, learning programs, and development workshops alongside industry professionals. Candidates may apply to a maximum of two positions across Our Company and its affiliates globally. Applications will be considered in the order they are submitted. Applications are reviewed on a rolling basis, so we encourage you to apply early. Please clearly state your availability in your resume, including your start and end dates.\n\nBuild evaluation systems for image and video models/agents, covering generation quality, instruction following, multimodal understanding, and safety.\nDevelop automated metrics and model-based evaluators, and design reproducible human evaluation protocols.\nDesign and develop video generation/debugging agents that orchestrate multi-step creative workflows.\nBuild large-scale image and video data pipelines, and turn evaluation findings into model and product improvements.\n\nQualifications\nMinimum Qualifications:\n\nCurrently pursuing a Bachelor's/ Master's degree in Computer Science, Artificial Intelligence, Machine Learning, Computer Vision, or a related field.\nSolid foundation in deep learning and computer vision, including generative modeling fundamentals.\nPractical experience in at least one of: visual generation, multimodal LLMs, video understanding, or visual quality assessment.\nStrong Python skills and proficiency with PyTorch or an equivalent framework, or multimodal evaluation framework.\nDemonstrated research or engineering ability through publications, substantial projects, internships, or open-source work.\n\nPreferred Qualifications:\n\nPublications at top‑tier vision or ML venues, e.g., NeurIPS, ICML, CVPR, ICCV, ECCV, etc.\nHands‑on experience with modern visual generation stacks, including diffusion-based models and their post‑training.\nFamiliarity with visual generation benchmarks, or experience building evaluation frameworks.\nExperience applying agent frameworks to creative workflows, or working with large‑scale video data infrastructure.\n\nJob Information\nAbout Us\nFounded in 2012, ByteDance's mission is to inspire creativity and enrich life. With a suite of more than a dozen products, including TikTok, Lemon8, CapCut and Pico as well as platforms specific to the China market, including Toutiao, Douyin, and Xigua, ByteDance has made it easier and more fun for people to connect with, consume, and create content.\nWhy Join ByteDance\nInspiring creativity is at the core of ByteDance's mission. Our innovative products are built to help people authentically express themselves, discover and connect – and our global, diverse teams make that possible. Together, we create value for our communities, inspire creativity and enrich life - a mission we work towards every day.\nAs ByteDancers, we strive to do great things with great people. We lead with curiosity, humility, and a desire to make impact in a rapidly growing tech company. By constantly iterating and fostering an \"Always Day 1\" mindset, we achieve meaningful breakthroughs for ourselves, our Company, and our users. When we create and grow together, the possibilities are limitless. Join us.\nDiversity & Inclusion\nByteDance is committed to creating an inclusive space where employees are valued for their skills, experiences, and unique perspectives. Our platform connects people from across the globe and so does our workplace. At ByteDance, our mission is to inspire creativity and enrich life. To achieve that goal, we are committed to celebrating our diverse voices and to creating an environment that reflects the many communities we reach. We are passionate about this and hope you are too.\nReasonable Accommodation\nByteDance is committed to providing reasonable accommodations in our recruitment processes for candidates with disabilities, pregnancy, sincerely held religious beliefs or other reasons protected by applicable laws. If you need assistance or a reasonable accommodation, please reach out to us at https://tinyurl.com/RA-request \n#J-18808-Ljbffr","company":"Pangle","rawCompany":"pangle","city":"Eastern","state":"KY","isRemote":false,"isActive":false,"createdAt":"2026-09-23T04:29:39.257Z","occupations":[{"code":"15-1221.00","title":"Computer and Information Research Scientists","slug":"computer-and-information-research-scientists"},{"code":"15-2051.00","title":"Data Scientists","slug":"data-scientists"},{"code":"15-1252.00","title":"Software Developers","slug":"software-developers"}],"industries":[{"code":"541511","title":"Custom Computer Programming Services","slug":"custom-computer-programming-services"},{"code":"541715","title":"Research and Development in the Physical, Engineering, and Life Sciences (except Nanotechnology and Biotechnology)","slug":"research-and-development-in-the-physical-engineering-and-life-sciences-except-nanotechnology-and-biotechnology"},{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"Visual Generation & Multimodal Evaluation Machine Learning Engineer Intern (AML-Ark-US) - 2027 [...]","description":"Join us as we work together to inspire creativity and enrich life around the globe.\nLocation\nSan Jose\nTeam\nTechnology\nEmployment Type\nIntern\nJob Code\nA174683\nShare this listing\nResponsibilities\nThe Applied Machine Learning Ark team combines system engineering and machine learning to develop and operate Large Language Model (LLM) service platforms that offer businesses Model-as-a-Service (MaaS) solutions, serving both large model providers and downstream users. The US team drives the design, development, and operation of MaaS solutions across the US and international markets outside mainland China. We are building full-stack, end-to-end solutions spanning text and multimodal LLM algorithms, LLM training/fine-tuning/inference frameworks, prompt engineering, model alignment, and intelligent agent systems. Beyond model serving, we operate large-scale log analytics pipelines that process massive volumes of invocation logs from text models, multimodal models, and agent systems — extracting usage patterns, quality signals, and actionable insights to inform model improvement, system optimization, and product decisions through continuous, data-driven feedback loops. We are actively seeking talented engineers and researchers specializing in Large Language Models and AI Agent systems to join our dynamic team. We are looking for talented individuals to join us for an internship. Our internship program offers students hands‑on experience, industry exposure, and opportunities to apply their knowledge to real‑world challenges while building a strong foundation for personal and professional growth. Interns will gain practical experience, explore potential career paths, and participate in social events, learning programs, and development workshops alongside industry professionals. Candidates may apply to a maximum of two positions across Our Company and its affiliates globally. Applications will be considered in the order they are submitted. Applications are reviewed on a rolling basis, so we encourage you to apply early. Please clearly state your availability in your resume, including your start and end dates.\n\nBuild evaluation systems for image and video models/agents, covering generation quality, instruction following, multimodal understanding, and safety.\nDevelop automated metrics and model-based evaluators, and design reproducible human evaluation protocols.\nDesign and develop video generation/debugging agents that orchestrate multi-step creative workflows.\nBuild large-scale image and video data pipelines, and turn evaluation findings into model and product improvements.\n\nQualifications\nMinimum Qualifications:\n\nCurrently pursuing a Bachelor's/ Master's degree in Computer Science, Artificial Intelligence, Machine Learning, Computer Vision, or a related field.\nSolid foundation in deep learning and computer vision, including generative modeling fundamentals.\nPractical experience in at least one of: visual generation, multimodal LLMs, video understanding, or visual quality assessment.\nStrong Python skills and proficiency with PyTorch or an equivalent framework, or multimodal evaluation framework.\nDemonstrated research or engineering ability through publications, substantial projects, internships, or open-source work.\n\nPreferred Qualifications:\n\nPublications at top‑tier vision or ML venues, e.g., NeurIPS, ICML, CVPR, ICCV, ECCV, etc.\nHands‑on experience with modern visual generation stacks, including diffusion-based models and their post‑training.\nFamiliarity with visual generation benchmarks, or experience building evaluation frameworks.\nExperience applying agent frameworks to creative workflows, or working with large‑scale video data infrastructure.\n\nJob Information\nAbout Us\nFounded in 2012, ByteDance's mission is to inspire creativity and enrich life. With a suite of more than a dozen products, including TikTok, Lemon8, CapCut and Pico as well as platforms specific to the China market, including Toutiao, Douyin, and Xigua, ByteDance has made it easier and more fun for people to connect with, consume, and create content.\nWhy Join ByteDance\nInspiring creativity is at the core of ByteDance's mission. Our innovative products are built to help people authentically express themselves, discover and connect – and our global, diverse teams make that possible. Together, we create value for our communities, inspire creativity and enrich life - a mission we work towards every day.\nAs ByteDancers, we strive to do great things with great people. We lead with curiosity, humility, and a desire to make impact in a rapidly growing tech company. By constantly iterating and fostering an \"Always Day 1\" mindset, we achieve meaningful breakthroughs for ourselves, our Company, and our users. When we create and grow together, the possibilities are limitless. Join us.\nDiversity & Inclusion\nByteDance is committed to creating an inclusive space where employees are valued for their skills, experiences, and unique perspectives. Our platform connects people from across the globe and so does our workplace. At ByteDance, our mission is to inspire creativity and enrich life. To achieve that goal, we are committed to celebrating our diverse voices and to creating an environment that reflects the many communities we reach. We are passionate about this and hope you are too.\nReasonable Accommodation\nByteDance is committed to providing reasonable accommodations in our recruitment processes for candidates with disabilities, pregnancy, sincerely held religious beliefs or other reasons protected by applicable laws. If you need assistance or a reasonable accommodation, please reach out to us at https://tinyurl.com/RA-request \n#J-18808-Ljbffr","datePosted":"2026-09-23T04:29:39.257Z","dateModified":"2026-09-23T04:29:39.257Z","hiringOrganization":{"@type":"Organization","name":"Pangle","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"Eastern","addressRegion":"KY","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"95495c24b80cc83cc0d3e30a"},"url":"https://jobsearcher.com/jobs/95495c24b80cc83cc0d3e30a"}}