About the role#
RunAnywhere is looking for an intern to help us grow our developer community. You will be the voice of our project, helping developers understand how to run AI models directly on mobile devices. You will work directly with our founders to build content and engage with users where they spend their time.
What you'll do#
- Write technical blogs and tutorials, including integration guides, benchmark reports, and explainers on topics like quantization and NPU performance.
- Create multimedia content such as demo videos, screen recordings, and simple sample apps to show our technology in action.
- Engage with developers on platforms like X, LinkedIn, Reddit, Hacker News, Discord, and GitHub to answer questions and join technical conversations.
- Support product and feature launches by drafting and distributing content.
- Use feedback from the community to identify and create new documentation, blog posts, or demos.
What you'll need#
- A working understanding of on-device AI inference, including concepts like quantization, GGUF, and inference engines.
- Strong written English skills and the ability to explain technical topics clearly.
- Experience creating your own content, such as demo videos, diagrams, or sample applications.
- Active participation on platforms like X or LinkedIn and comfort engaging with developers in technical discussions.
- Hands-on experience running local models on your own devices or using inference engines like llama.cpp.
- Familiarity with the local AI community, such as r/LocalLLaMA, Hugging Face, or Hacker News.
- Mobile development experience in Android, Kotlin, iOS, Swift, Flutter, or React Native is a plus.
- Prior experience in DevRel, open-source contributions, or technical blogging is a plus.
Location & details#
- This is a fully remote role.
- The position is paid at $750 to $1,500 per month, depending on experience and output.
- This is a full-time internship with flexible hours, though some overlap with the team's working hours is expected.
- The term is rolling.
About RunAnywhere (YC W26)
RunAnywhere provides software for running open-weight models across various environments. The company develops the Wally inference stack to support serverless, bring-your-own-cloud, and on-premise deployments. It also maintains open-source SDKs for local execution on mobile and desktop operating systems. Founded in 2025, the firm is based in San Francisco and employs 9 people.
How to get in at RunAnywhere (YC W26)
Securing an internship at RunAnywhere requires speed, as early applicants are often reviewed before the candidate pool becomes unmanageable. Intern Insider sends an instant alert the moment a role matching your target is published, helping you apply among the first to ensure your materials are seen. This proactive approach helps you avoid the frustration of submitting applications into a growing, unseen pile. You can also use Intern Insider to surface the recruiters behind RunAnywhere roles to reach out directly. Connecting with a recruiter to ask about the team or a referral often improves your response rate compared to standard portal submissions. It is a practical way to show genuine interest while navigating the competitive landscape of early-stage software companies.



