LLaSM
PulseGate's liveness check found it on 3 Oct 2026; it is registered on Hugging Face Spaces and has been in the index since 31 Jul 2026. How this is checked
LLaSM is a demo of a large language-speech model that combines speech recognition, language understanding, and text-to-speech in one system. Users can type or speak into the interface; the model transcribes, reasons about the request, and replies with synthesized speech. This end-to-end voice interface showcases advances in multimodal language models for more natural human-AI interaction.
Inferred · not functionally tested
Overview
3 featuresPurpose: Enabling natural voice-based conversations with large language models without separate transcription and synthesis steps.
Inferred · not functionally tested
Audience: AI researchers and voice interface developers
Inferred · not functionally tested
Functions: speech_to_text, voice_synthesis
Inferred · not functionally tested
Interfaces: API: unknown · MCP: unknown · CLI: unknown · Self-hosting: unknown
Recorded constraints: pricing: free · license: Proprietary · platforms: WEB · deployment: browser, cloud_managed
Constraint provenance is unknown; confirm requirements with the publisher.
Record sources: huggingface.co. These links do not verify the individual claims.
LLaSM sits in PulseGate's Other voice & speech category. Inferred · not functionally tested: It focuses on enabling natural voice-based conversations with large language models without separate transcription and synthesis steps. Inferred · not functionally tested: It is built as an open-source project for AI researchers and voice interface developers. Basis unknown · not verified: It is available for free. Basis unknown · not verified: It runs on the web.
Behind LLaSM is LinkSoul. Inferred · not functionally tested: Among its 3 catalogued features are Voice Input, Speech Output, and Multimodal Chat.
Summary written by a language model from the project’s public pages.
Tasks: Inferred · not functionally tested
- Voice Input
- Speech Output
- Multimodal Chat
Topics: Inferred · not functionally tested
Built with & integrations
- AWS
- x-amz-cf-id header · x-amz-cf-pop header · via header
- local_oss
- huggingface in the HTML
- meta_llama
- llama- in the HTML
Trust & compliance
Indexing history
1What PulseGate has recorded for this listing
- Indexed31 Jul · 02:35 UTCLinkSoul/LLaSM seen via Hugging Face EnumeratorSource: Hugging Face Enumerator · Open
Frequently asked questions about LLaSM
- What does LLaSM do?
- Inferred · not functionally tested: LLaSM focuses on enabling natural voice-based conversations with large language models without separate transcription and synthesis steps. It is catalogued under Other voice & speech on PulseGate.
- Who should use LLaSM?
- Inferred · not functionally tested: LLaSM is an open-source project built for AI researchers and voice interface developers.
- Is LLaSM free?
- Basis unknown · not verified: Yes — LLaSM is free to use.
- What platforms does LLaSM run on?
- Basis unknown · not verified: LLaSM runs on the web.
- Is LLaSM still active?
- PulseGate's liveness check found it on 3 Oct 2026.
- What are alternatives to LLaSM?
- Similar projects tracked by PulseGate include LLaDA, Video LLaMA, and Chinese LLaVA.LLaDAVideo LLaMAChinese LLaVA
- Who develops LLaSM?
- LLaSM is developed by LinkSoul.
Similar projects
Closest matches by what these projects do