Skip to content
View vigorlee's full-sized avatar
🎯
Focusing
🎯
Focusing

Highlights

  • Pro

Block or report vigorlee

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
vigorlee/README.md

Mingyi Li

Embodied AI · Multimodal Reasoning · Robot Navigation

I build embodied AI systems that turn multimodal perception into grounded robot decisions through structured memory, retrieval, world models, and verified execution. I am with Beijing Institute of Technology.

Current focus

  • Knowledge-grounded intelligence — multimodal retrieval, semantic-spatial memory, and open-world ObjectNav
  • Reliable autonomy — world-model-assisted control, safety checks, and multimodal terminal verification
  • Reproducible robotics — inspectable simulation assets, scripted validation, and clear system boundaries

Selected projects

An embodied multimodal retrieval prototype for open-world object-goal navigation. It connects visual observations, spatial-region memory, semantic evidence, and vision-language model backends, with a path toward simulation-to-real deployment.

A map-independent charging workflow for Go2-W that combines world-model-generated action chunks, cross-embodiment motion adaptation, veto-only safety checks, and multimodal terminal verification.

A reproducible Isaac Sim scene with scripted setup, checksum verification, headless rendering and physics checks, and explicit boundaries for third-party assets.

Research interests

Embodied AI · Object-goal navigation · Multimodal RAG · Vision-language models · World models · Robot learning · Isaac Sim

Working principles

I care about clear system boundaries, reproducible setup, and evidence that others can inspect. I try to document both what a prototype demonstrates and what it does not.

Contact

Popular repositories Loading

  1. Multimodal--RAG Multimodal--RAG Public

    EMKG: multimodal memory and retrieval for open-world object-goal navigation in dynamic environments.

    Python 51

  2. RoboKino RoboKino Public

    3

  3. wave-go wave-go Public

    Successful Go2-W mapless charging demo backup

    Python 1

  4. SmarterStreaming SmarterStreaming Public

    Forked from daniulive/SmarterStreaming

    国内外为数不多致力于极致体验的超强全自研跨平台(windows/android/iOS)流媒体内核,通过模块化自由组合,支持实时RTMP推流、RTSP推流、RTMP播放器、RTSP播放器、录像、多路流媒体转发、音视频导播、动态视频合成、音频混音、直播互动、内置轻量级RTSP服务等,比快更快,业界真正靠谱的超低延迟直播SDK(1秒内,低延迟模式下200~400ms)。

    Java

  5. EasyPlayer-RTSP-Android EasyPlayer-RTSP-Android Public

    Forked from yiqideren/EasyPlayer_Android

    An elegant, simple, fast android RTSP/RTMP/HLS/HTTP Player.EasyPlayer support RTSP(RTP over TCP/UDP)version & Pro version,cover all kinds of streaming media!EasyPlayer是一款精炼、高效、稳定的流媒体播放器,分为RTSP版和Pro…

    Java

  6. TensorFlow2.0-Examples TensorFlow2.0-Examples Public

    Forked from DjangoPeng/TensorFlow2.0-Examples

    🙄 difficult algorithm, simple code.

    Jupyter Notebook