Description
實習目標
加入 Vulcan 團隊,與研究員及工程師並肩作戰!你將參與守護大型語言模型(LLM)安全的前線工作,深入探索 AI 漏洞檢測、數據隱私保護及模型防禦機制,累積最紮實的 AI 安全實務經驗。
核心職責
我們將複雜的技術開發轉化為具體的實習任務,讓你快速上手:
- 防護機制驗證數據生成 (DLP & Moderation):
- 任務內容:設計用來測試「防禦牆」的輸入資料,模擬各類異常行為。
- 實作重點:撰寫測試案例,驗證藍隊的數據外洩防護(DLP)與內容審核(Moderation)機制是否能精準攔截個資外洩、暴力或歧視等不當言論。
- 數據標籤與分類 (Labeling):
- 任務內容:擔任 AI 的「安全導師」,定義行為邊界。
- 實作重點:精準標註訓練與測試資料,識別「隱私洩漏」或「提示詞劫持(Prompt Injection)」等惡意指令,協助模型學習辨識風險。
- 測試結果分析與洞察:
- 任務內容:挖掘 AI 的潛在弱點,將數據轉化為情報。
- 實作重點:整理 Vulcan 平台的測試結果,分析模型在不同安全維度的表現,並協助團隊製作直觀的分析圖表與技術報告。
上班時間 09:00-18:00 (每月 3-5天,依工作量彈性安排)
Internship Goal
Join the Vulcan team and work alongside top-tier researchers and engineers to fortify the security of Large Language Models (LLMs). This role offers a unique, hands-on opportunity to dive deep into AI vulnerability detection, data privacy, and model defense strategies.
Key Responsibilities
We have streamlined our core technical workflows into actionable tasks for you to hit the ground running:
- Defense Mechanism Validation (DLP & Moderation):
- What you’ll do: Generate test data designed to challenge and verify our "defense walls" by simulating various edge cases.
- Execution: Craft test cases to evaluate whether Blue Team defenses, such as Data Loss Prevention (DLP) and Content Moderation, can accurately intercept sensitive info leaks or harmful content, identifying gaps in the current security layer.
- Data Labeling & Classification:
- What you’ll do: Act as a "Safety Mentor" for the AI, teaching it to distinguish between safe and hazardous interactions.
- Execution: Annotate training and testing datasets by identifying "Privacy Leaks" or "Prompt Injection" attacks—malicious instructions designed to bypass AI safety guardrails.
- Testing Analysis & Reporting:
- What you’ll do: Uncover AI vulnerabilities and document key findings.
- Execution: Organize testing data from the Vulcan platform, identify weak spots across various security dimensions, and assist the team in transforming these insights into clear charts and professional reports.
Interview Process
- Phone interview (30mins)
- Interview (1hr) : 45 mins with Hiring Managers, 15 mins with HR
Similar jobs
想要讓你的語言能力跟上 AI 浪潮嗎? 我們正在找尋對文字敏銳、對 AI 有好奇心的夥伴。這份工作不需要你會寫程式,而是需要你用中文、英文、韓文或阿拉伯文,幫忙檢測 AI 在不同語言環境下夠不夠安全、會不會被「教壞」。 如果你平常就在用 AI 工具,且喜歡跟機器人對話,你就是我們要找的人! 🛠️ 工作內容: 我們把複雜的技術簡化為兩個重點任務: 當 AI 的「行為導師」(數據標註與分類): AI 就像小孩,需要有人告訴它什麼能說、什麼…
Est. 120,000 USD
About Us AIFT is at the forefront of AI safety and security, developing cutting-edge solutions, Vulcan Attack (automated vulnerability assessment/red teaming) and Vulcan Protect (real-time monitoring), to ensure AI syste…
Job Overview We are seeking a skilled, hands-on Senior Forward Deployed Engineer (FDE) with a strong technical background and exceptional communication skills. Standing at the front lines of customer engagement, you will…
Job Overview We are seeking a skilled and passionate Site Reliability Engineer with a strong technical background and excellent communication skills. This individual will lead the development, construction, and managemen…
Job Overview OneSavie Lab是一個結合 Web3與AI科技的頂尖實驗室,致力於為Web3 生態系統以及眾多參與者提供先進的智能風險監控和管理工具,包括中心化和去中心化的多元策略選擇。 OneSavie Lab的核心使命是為投資者、機構和其他虛擬資產領域的利益相關者提供全方位的支持。從代幣風險監控、資產情報分析,到Web3的風控諮詢,我們都能提供專業、精確的服務,在新興領域為客戶護航。 OneSavieLab 在招募…
Job Overview OneSavie Lab, OneInfinity 是一個結合 Web3與AI科技的頂尖實驗室,致力於為Web3 生態系統以及眾多參與者提供先進的智能風險監控和管理工具,包括中心化和去中心化的多元策略選擇。 OneSavie Lab的核心使命是為投資者、機構和其他虛擬資產領域的利益相關者提供全方位的支持。從代幣風險監控、資產情報分析,到Web3的風控諮詢,我們都能提供專業、精確的服務,在新興領域為客戶護航。 O…
We are looking for a highly skilled AI Engineer specializing in Large Language Models (LLMs) and Agentic AI. You will architect, build, and deploy production-grade LLM applications — from intelligent knowledge bases and…
Job Overview OneSavie Lab, OneInfinity 是一個結合 Web3與AI科技的頂尖實驗室,致力於為Web3 生態系統以及眾多參與者提供先進的智能風險監控和管理工具,包括中心化和去中心化的多元策略選擇。 OneSavie Lab的核心使命是為投資者、機構和其他虛擬資產領域的利益相關者提供全方位的支持。從代幣風險監控、資產情報分析,到Web3的風控諮詢,我們都能提供專業、精確的服務,在新興領域為客戶護航。 O…
Est. 418,650 USD
Veeam is the Data and AI Trust Company, specializing in helping organizations ensure their data and AI are fully understood, secured, and resilient to enable the acceleration of safe AI at scale. As the market leader in…
SonicWall is a cybersecurity forerunner with more than 30 years of expertise and is recognized as a leading partner-first company, ensuring our partners and their customers are never alone in the fight against cybercrime…
Est. 141,000 USD
ABOUT LVT LVT is redefining how businesses operate in the physical world, moving beyond traditional security solutions to deliver AI-driven, actionable intelligence that makes sites smarter, safer, and more secure. Since…
Est. 208,725 USD
ABOUT LVT LVT is redefining how businesses operate in the physical world, moving beyond traditional security solutions to deliver AI-driven, actionable intelligence that makes sites smarter, safer, and more secure. Since…
Est. 80,000 USD
About Workato Workato delivers enterprise infrastructure for the agentic era, redefining iPaaS and helping enterprises unify data, applications, processes, and AI into a single, governed platform. A leader in Enterprise…
Est. 120,000 PLN
Xebia is a global AI-first, digital transformation, and engineering partner. With over 25 years of experience and a team of 5,000 professionals across 16 countries, we help organizations design and build scalable product…
Est. 84,000 EUR
This role is located in Budapest, Hungary - We are a hybrid work environment and are in the office 3+ days/per week. Tulip, the leader in AI-native frontline operations, is helping companies around the world equip their…
Cymetrics 是亞洲領先的資安原廠之一,擁有專屬的高端資安產品。我們提供專業的紅隊演練、滲透測試和弱點掃描服務,集結了工程技術與資安專長的團隊。 團隊成員均擁有資安風險管理和滲透測試的專業知識,具有在四大管理顧問公司、領導資安服務商、知名品牌原廠的豐富經驗,且積極參與國際 CTF 競賽,並曾取得世界第三名。我們服務的客戶來自不同產業,包括政府、金融、製造業、高科技和電子商務等等。我們團隊也協助集團獲得 ISO 27001 和 IS…
Est. 273,875 USD
XPENG is a leading smart technology company at the forefront of innovation, integrating advanced AI and autonomous driving technologies into its vehicles, including electric vehicles (EVs), electric vertical take-off and…
Est. 60,000 USD
Who We Are VML is a leading creative company that combines brand experience, customer experience, and commerce, creating connected brands to drive growth. VML is celebrated for its innovative and award-winning human-firs…
About Workato Workato delivers enterprise infrastructure for the agentic era, redefining iPaaS and helping enterprises unify data, applications, processes, and AI into a single, governed platform. A leader in Enterprise…
Est. 140,000 USD
Location: United States – RemoteClearance: Ability to obtain and maintain a Public Trust LTS is seeking a highly skilled Senior Applied AI Engineer to focus on continuously improving the intelligence behind the platform.…
Est. 196,450 USD
ABOUT LVT LVT is redefining how businesses operate in the physical world, moving beyond traditional security solutions to deliver AI-driven, actionable intelligence that makes sites smarter, safer, and more secure. Since…
SonicWall is a cybersecurity forerunner with more than 30 years of expertise and is recognized as a leading partner-first company, ensuring our partners and their customers are never alone in the fight against cybercrime…
Who We Are VML is a leading creative company that combines brand experience, customer experience, and commerce, creating connected brands to drive growth. VML is celebrated for its innovative and award-winning human-firs…
About the role We are seeking an experienced Machine Learning Lead to helm our Machine Learning team. In this pivotal role, you will be the engineering architect behind Vulcan’s core AI capabilities. You will act as the…
Est. 120,000 GBP
Isomorphic Labs is applying frontier AI to help unlock deeper scientific insights, faster breakthroughs, and life-changing medicines with an ambition to solve all disease. The future is coming. A future enabled and enric…
Est. 296,450 USD
Vectra® is the leader in AI-driven threat detection and response for hybrid and multi-cloud enterprises. The Vectra AI Platform delivers integrated signal across public cloud, SaaS, identity, and data center networks in…
Est. 140,000 USD
LTS is seeking an AI Platform and Harness Engineer to develop and maintain the infrastructure, tooling, and evaluation frameworks that power enterprise AI solutions. This role is responsible for building the AI platform…
Est. 620,000 USD
About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of co…
OPSWAT, a global leader in IT, OT, and ICS critical infrastructure cybersecurity, delivers an end-to-end platform that gives public and private sector organizations and enterprises the critical advantage needed to protec…
Est. 90,000 USD
We are searching for an AI Safety Specialist who will play a crucial role in enhancing the security and robustness of language models. You will ensure the safe deployment of AI systems by conducting adversarial testing,…