Inference Providers
Active filters: web-llm
welcoma/gemma-4-E2B-it-q4f16_1-MLC
Text Generation
• Updated • 3
mlc-ai/Llama-2-7b-chat-hf-q4f16_1-MLC
Updated • 806
• 9
mlc-ai/Llama-2-7b-chat-hf-q4f32_1-MLC
Updated • 50
• 2
mlc-ai/Mistral-7B-Instruct-v0.2-q4f16_1-MLC
Updated • 992
• 4
mlc-ai/Llama-2-13b-chat-hf-q4f16_1-MLC
Updated • 117
• 1
mlc-ai/Llama-2-70b-chat-hf-q4f16_1-MLC
Updated • 3
• 2
mlc-ai/OpenHermes-2.5-Mistral-7B-q4f16_1-MLC
Updated • 83
• 1
mlc-ai/NeuralHermes-2.5-Mistral-7B-q4f16_1-MLC
Updated • 51
• 2
mlc-ai/WizardMath-7B-V1.0-q4f16_1-MLC
mlc-ai/WizardMath-13B-V1.0-q4f16_1-MLC
mlc-ai/WizardMath-70B-V1.0-q4f16_1-MLC
mlc-ai/RedPajama-INCITE-Chat-3B-v1-q4f16_1-MLC
Updated • 941
• 3
mlc-ai/RedPajama-INCITE-Chat-3B-v1-q4f32_1-MLC
Updated • 43
• 2
mlc-ai/WizardMath-7B-V1.1-q4f16_1-MLC
Updated • 801
• 2
mlc-ai/phi-1_5-q4f16_1-MLC
Updated • 195
Updated • 6
• 1
mlc-ai/gpt2-medium-q0f16-MLC
Updated • 7
• 1
mlc-ai/Mistral-7B-Instruct-v0.2-q3f16_1-MLC
Updated • 13
• 4
mlc-ai/TinyLlama-1.1B-Chat-v0.4-q4f16_1-MLC
Updated • 306
• 1
mlc-ai/TinyLlama-1.1B-Chat-v0.4-q4f32_1-MLC
Updated • 294
• 1
mlc-ai/TinyLlama-1.1B-Chat-v0.4-q0f16-MLC
mlc-ai/TinyLlama-1.1B-Chat-v0.4-q0f32-MLC
mlc-ai/phi-1_5-q4f32_1-MLC
Updated • 59
• 1
Updated • 6
• 2
mlc-ai/NeuralHermes-2.5-Mistral-7B-q3f16_1-MLC