This repo contains the original implementation of VAuLT, the Vision-and-Augmented-Language Transformer. We provide instructions to download some multimodal social-media datasets, and scripts to experiment with. VAuLT is a stack of Transformers, a LM like BERT that preprocesses the text input of ViLT
A Comprehensive Speech Processing Algorithms Library for research and production use
DiscordNPC lets you interact with ChatGPT through a Discord voice channel, enabling a natural conversation.
Listen4Me bot converts audio messages into text.
An OpenAI Compatible API which integrates LLM, Embedding and Reranker. 一个集成 LLM、Embedding 和 Reranker 的 OpenAI 兼容 API
Drop-in OIDC & Google A2A auth + Weaviate memory for Ollama, vLLM and any local LLM server.
Calibrating LLMs with Information-Theoretic Evidential Deep Learning (ICLR 2025)
Multiple Choice Learning of Low Rank Adapters for Language Modeling
Multi-scale Finetuning for Encoder-based Time Series Foundation Models
an easy way to create JSONL files for fine-tuning openai models.
Sparse Autoencoders (SAE) vs CLIP fine-tuning fun.
共 7255 条 · 第 341 / 363 页