TGViewer
Channel Public Channel
Github LLMs

Github LLMs

@llm_learning

Subscribers
756
Photos
40
Videos
3
Links
55

Showing posts older than #26 Β· Back to latest

Older Posts 20 shown
Post #22 950

Forwarded from Machine learning books and papers

🌟 BioNeMo: A Framework for Developing AI Models for Drug Design.

NVIDIA BioNeMo2 Framework is a set of tools, libraries, and models for computational drug discovery and design.



▢️ Pre-trained models:

🟒 ESM-2 is a pre-trained bidirectional encoder (BERT-like) for amino acid sequences. BioNeMo2 includes checkpoints with parameters 650M and 3B;

🟒 Geneformer is a tabular scoring model that generates a dense representation of a cell's scRNA by examining co-expression patterns in individual cells.


▢️ Datasets:

🟠 CELLxGENE is a collection of publicly available single-cell datasets collected by the CZI (Chan Zuckerberg Initiative) with a total volume of 24 million cells;


🟠 UniProt is a database of clustered sets of protein sequences from UniProtKB, created on the basis of translated genomic data.



🟑 Project page
🟑 Documentation
πŸ–₯ GitHub

@Machine_learn
  • πŸ‘ 2
Post #19 969

Forwarded from Machine learning books and papers

⚑️ MobileLLM


🟒MobileLLM-125M. 30 Layers, 9 Attention Heads, 3 KV Heads. 576 Token Dimension;

🟒MobileLLM-350M. 32 Layers, 15 Attention Heads, 5 KV Heads. 960 Token Dimension;

🟒MobileLLM-600M. 40 Layers, 18 Attention Heads, 6 KV Heads. 1152 Token Dimension;

🟒MobileLLM-1B. 54 Layers, 20 Attention Heads, 5 KV Heads. 1280 Token Dimension;


🟑Arxiv
πŸ–₯GitHub


@Machine_learn
Post #15 4.23K
πŸ“– LLM-Agent-Paper-List is a repository of papers on the topic of agents based on large language models (LLM)! The papers are divided into categories such as LLM agent architectures, autonomous LLM agents, reinforcement learning (RL), natural language processing methods, multimodal approaches and tools for developing LLM agents, and more.

πŸ–₯ Github

https://t.me/deep_learning_proj
  • πŸ‘ 3
Post #14 3.62K
🌟 Zamba2-Instruct

Π’ сСмСйствС 2 ΠΌΠΎΠ΄Π΅Π»ΠΈ:

🟒Zamba2-1.2B-instruct;
🟠Zamba2-2.7B-instruct.



# Clone repo
git clone https://github.com/Zyphra/transformers_zamba2.git
cd transformers_zamba2

# Install the repository & accelerate:
pip install -e .
pip install accelerate

# Inference:
from transformers import AutoTokenizer, AutoModelForCausalLM
import torch

tokenizer = AutoTokenizer.from_pretrained("Zyphra/Zamba2-2.7B-instruct")
model = AutoModelForCausalLM.from_pretrained("Zyphra/Zamba2-2.7B-instruct", device_map="cuda", torch_dtype=torch.bfloat16)

user_turn_1 = "user_prompt1."
assistant_turn_1 = "assistant_prompt."
user_turn_2 = "user_prompt2."
sample = [{'role': 'user', 'content': user_turn_1}, {'role': 'assistant', 'content': assistant_turn_1}, {'role': 'user', 'content': user_turn_2}]
chat_sample = tokenizer.apply_chat_template(sample, tokenize=False)

input_ids = tokenizer(chat_sample, return_tensors='pt', add_special_tokens=False).to("cuda")
outputs = model.generate(**input_ids, max_new_tokens=150, return_dict_in_generate=False, output_scores=False, use_cache=True, num_beams=1, do_sample=False)
print((tokenizer.decode(outputs[0])))





πŸ–₯GitHub

https://t.me/deep_learning_proj
  • πŸ‘ 2
Post #12 4.41K
Post #10 3.18K
🌟 GRIN MoE: Mixture-of-Experts ΠΎΡ‚ Microsoft.


🟒total parameters: 16x3.8B;
🟒active parameters: 6.6B;
🟒context length: 4096;
🟒number of embeddings 4096;
🟒number of layers: 32;
βœ…https://t.me/deep_learning_proj


🟑Arxiv
🟑Demo
πŸ–₯Github
Post #9 4.73K
Post #8 601
MiniCPM-V

MiniCPM-V 2.6: A GPT-4V Level MLLM for Single Image, Multi Image and Video on Your Phone
                                                                   
Creator: OpenBMB
Stars ⭐️: 11.4k
Forked By: 798
GitHub Repo:
https://github.com/OpenBMB/MiniCPM-V

βž–βž–βž–βž–βž–βž–βž–βž–βž–βž–βž–βž–βž–βž–       
Join βœ…https://t.me/deep_learning_proj
GitHub GitHub - OpenBMB/MiniCPM-V: A Pocket-Sized MLLM for Ultra-Efficient Image and Video Understanding on Your Phone A Pocket-Sized MLLM for Ultra-Efficient Image and Video Understanding on Your Phone - OpenBMB/MiniCPM-V
Post #3 592
firecrawl

Turn entire websites into LLM-ready markdown or structured data. Scrape, crawl and extract with a single API.
                                                                   
Creator: Mendable
Stars ⭐️: 12.3k
Forked By: 861
GitHub Repo:
https://github.com/mendableai/firecrawl

βœ… https://t.me/deep_learning_proj
GitHub GitHub - firecrawl/firecrawl: Supercharge your AI agents with data from the web and beyond. Building the library for superintelligence.… Supercharge your AI agents with data from the web and beyond. Building the library for superintelligence. πŸ”₯ - firecrawl/firecrawl
  • πŸ‘ 2
Post #2 501
graphrag

A modular graph-based Retrieval-Augmented Generation (RAG) system
                                                                   
Creator: Microsoft
Stars ⭐️: 13.7k
Forked By: 1.2k
GitHub Repo:
https://github.com/microsoft/graphrag

βž–βž–βž–βž–βž–βž–βž–βž–βž–βž–βž–βž–βž–βž–       
Join @deep_learning_proj
GitHub GitHub - microsoft/graphrag: A modular graph-based Retrieval-Augmented Generation (RAG) system A modular graph-based Retrieval-Augmented Generation (RAG) system - microsoft/graphrag
  • πŸ‘ 1
Older posts β†’
Threads Profile ViewerView any public Threads profile without an account.Open ThreadLook β†’Writing with AI? Make it sound human.Metric37 rewrites AI drafts so they read naturally. Free AI detector, 1,500 words free.Try Metric37 β†’