PUMA Vault

Etiqueta: quantization

4 artículos con esta etiqueta.

  • 15 jun 2026

    Natural Language Processing with Transformers: Building Language Applications with Hugging Face

    • literature
    • transformers
    • hugging-face
    • nlp
    • fine-tuning
    • quantization
    • ollama
    • open-weights
    • mistral
    • gemma
    • bert
    • puma-core
    • book
    • implementation
    • python
    • literature-note
    • moc
  • 15 jun 2026

    Fine-Tuning LLMs — LoRA, QLoRA, GGUF Quantization, and PUMA Considerations

    • permanent
    • fine-tuning
    • lora
    • qlora
    • quantization
    • gguf
    • ollama
    • llm
    • training
    • puma-core
    • research
    • agents
    • benchmark
    • local-models
    • effort-estimation
    • issue-triage
    • architecture
  • 15 jun 2026

    LLM Models Used in PUMA — Technical Reference

    • permanent
    • llm
    • models
    • llama
    • mistral
    • phi
    • gemma
    • deepseek
    • gpt4o
    • qwen
    • moe
    • quantization
    • ollama
    • puma-core
    • research
    • benchmark
    • effort-estimation
    • issue-triage
    • agents
    • architecture
  • 15 jun 2026

    PN — Inference runs locally via a model server exposing an OpenAI-compatible…

    • permanent-note
    • puma
    • methodology
    • metrics
    • quantization

Creado con Quartz v4.5.2 © 2026

  • GitHub
  • Discord Community