G2继续向下滚动KEEP SCROLLING
H7ECOLOGICAL DIS-INTEGRATION
Water Treatment Plant Design: An Ecological Mode for Singapore, Fei Xiong Portfolio
H8DECODER-ONLY TRANSFORMER
A Vision-Language Model, after Vaswani 2017 & Fuyu 2023
A1熊非
P1

我是熊非,同济建筑学士学位Harvard 21'(MDes RR首个大陆录取本科生)。曾任豆包多模态产品负责人(包括图片/视频理解,实时视频通话,computer use);即梦founding PM;连续创业者,曾悉空科技COO(Spatial Intelligence)、曾画龙CEO(物理世界定制AI)。

2019在Harvard和MIT开始我的Computer Vision旅程,而后我致力于打造以视觉为统一接口的交互新范式,以及下一代视觉交互的前沿模型。我相信模型作为新质生产力,下一代的生产关系会围绕模型展开,最强的学习信号会来自产品落地和部署环境,有竞争力的组织会从这样的新生产力和生产关系长出来。我希望建立这样的组织,也希望认识更多有趣的朋友。

I'm Fei Xiong — B.Arch from Tongji, Harvard '21 (first mainland undergrad admitted to MDes RR). Formerly Head of Multimodal Product at Doubao (image/video understanding, real-time video calls, computer use); founding PM of Dreamina (ByteDance); serial entrepreneur — COO of PlaceInt (spatial intelligence), CEO of LongAI (custom AI for the physical world).

My computer vision journey began at Harvard and MIT in 2019. Since then I've focused on one thing: making vision the universal interface between humans and AI, and building the frontier models behind it. I believe models are the new means of production: the next production relations will form around them, the strongest learning signal will come from shipped products and deployment, and the most competitive organizations will grow out of both. That's the organization I want to build — and I'd love to meet more interesting people along the way.

2026

E1豆包,Computer Use Agent负责人Doubao · Head of Computer Use Agent

我负责豆包CUA的产品,新增【办公任务】大入口,设计GUI/非GUI混合策略,本地优先I lead Doubao's CUA product — launched the Office Tasks entry, designed the hybrid GUI/non-GUI execution strategy, local-first.

2024 — 25

E2豆包 ,VLM多模态交互负责人Doubao · Head of VLM Multimodal Interaction

我从零打造豆包图片理解,长视频理解,视频实时通话等功能入口;贡献X千万DAUI built Doubao's image understanding, video understanding and real-time video calls from scratch; contributed tens of millions of DAU.

2023 — 24

E3剪映/即梦,Smart EditCapCut / Dreamina · Smart Edit

我打造了业界首个Agent创作产品Smart Edit,后改名为小云雀/剪小映We built Smart Edit, the industry's first agentic creation product — later relaunched as Dreamina's AI-creation agents, Xiaoyunque (Skylark) / Jianxiaoying.

2023

E6画龙 LongAI · 创始人 / CEOLongAI · Founder / CEO

我打造了一个物理世界的 AI 私人定制设计平台,基于POD,让普通人设计定制,支持500+ SKUI built a custom-design AI platform for the physical world — print-on-demand, 500+ SKUs, helping people design their own products.

2021 — 23

E4龙湖集团 · CV 产品负责人Longfor Group · Head of CV Product

我负责在全球1000+自持物理空间,大规模边缘部署算力,涵盖住宅/售楼处/购物中心/后勤等We deployed edge AI at scale across 1,000+ self-owned properties worldwide — residences, sales galleries, malls and logistics.

2019 — 21

E7深空科技 PlaceInt · 联合创始人 / COOPlaceInt · Co-founder / COO

我基于目标检测/跟踪,reID等CV技术,构建空间智能体系,用端侧算力捕捉人群行为。服务20+大型空间,1000+摄像头,45 万㎡;团队有多篇NeurIPS / IEEE;客户包括香港置地、富力、合生创展We built a spatial-intelligence stack on detection, tracking and re-ID, capturing crowd behavior with edge compute. 20+ large venues, 1,000+ cameras, 450,000 m²; multiple NeurIPS / IEEE papers on the team; clients incl. Hongkong Land, R&F, Hopson.

2018 — 21

E11MIT · CSAIL/Media Lab/Senseable City LabMIT · CSAIL / Media Lab / Senseable City Lab

Cross Registration,主要为Course 6系列(CS,ML) 和Computer Vision系列Cross-registration — mainly Course 6 (CS, ML) and computer vision.

2018 — 21

E8Harvard GSD

GSD Mdes RR专业,第一个录取大陆本科生。MIT联合培养MDes RR — first mainland undergraduate admitted; jointly trained with MIT.

2015

E12NUS CDE

新加坡国立大学交换半年,方向:建筑技术Exchange semester at the National University of Singapore, focusing on architecture technology.

2013 — 18

E9同济大学 · 建筑学 B.ArchTongji University · B.Arch

建筑设计专业Architectural design.

W1豆包 · Computer UseDoubao · Computer Use增加办公任务大入口,设计GUI/非GUI混合执行策略,本地优先。Launched the Office Tasks entry; hybrid GUI/non-GUI execution, local-first.微信推文 ↗WECHAT ↗W2豆包 · 图片理解Doubao · Image Understanding从零构建豆包的视觉能力,增加图片/相册入口,从 0 到千万级 DAU。Built Doubao's visual capability from scratch — image & album entries, zero to tens of millions of DAU.微信推文 ↗WECHAT ↗W3豆包 · 视频通话Doubao · Video Call实时交互视频通话,豆包评穿搭/看小孩写作业等传播度广,边看边聊的多模态交互。Real-time interactive video calls — outfit checks, homework help: see-while-you-chat multimodal interaction.微信推文 ↗WECHAT ↗W4Bear 的本质学Bear's Essentialism用第一性原理拆解世界的长期专栏。A long-form column dissecting the world from first principles.飞书文档 ↗FEISHU DOC ↗W5Fundamentals: AI 播客Fundamentals: AI Podcast全自动生产的 AI 播客:海外长文→双角色逐字稿→TTS→自动上架小宇宙,每天更新。A fully automated AI podcast — long-form articles to two-host scripts to TTS, auto-published daily.飞书文档 ↗FEISHU DOC ↗W6画龙 LongAILongAI物理世界的 AI 私人定制设计平台,让普通人用AI创造品牌。A custom-design AI platform for the physical world — anyone can build a brand with AI.微信推文 ↗WECHAT ↗W7Place.Int 深空科技Place.Int基于计算机视觉的大型空间人群分析引擎,1000+ 摄像头规模落地。A computer-vision crowd-analytics engine for large venues — 1,000+ cameras in production.微信推文 ↗WECHAT ↗