Skip to content
View michaelnny's full-sized avatar
Block or Report

Block or report michaelnny

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Please don't include any personal information such as legal names or email addresses. Maximum 100 characters, markdown supported. This note will be visible to only you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse

Pinned

  1. InstructLLaMA InstructLLaMA Public archive

    Implements pre-training, supervised fine-tuning (SFT), and reinforcement learning from human feedback (RLHF), to train and fine-tune the LLaMA2 model to follow human instructions, similar to Instru…

    Jupyter Notebook 33 9

  2. MM-LLaMA MM-LLaMA Public archive

    Bring multimodality to the LLaMA model by leveraging ImageBind as the modal encoder. This project supports vision input (both images and short videos) to the LLaMA model, with text output generated…

    Python 3 2

  3. RAG-LLaMA RAG-LLaMA Public archive

    A clean and simple implementation of Retrieval Augmented Generation (RAG) to enhanced LLaMA chat model to answer questions from a private knowledge base. We use Tesla user manuals to build the know…

    Jupyter Notebook 2

  4. alpha_zero alpha_zero Public archive

    A PyTorch implementation of DeepMind's AlphaZero agent to play Go and Gomoku board games

    Python 44 11

  5. deep_rl_zoo deep_rl_zoo Public archive

    A collection of Deep Reinforcement Learning algorithms implemented with PyTorch to solve Atari games and classic control tasks like CartPole, LunarLander, and MountainCar.

    Python 94 7