Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up

AI Safety & Interpretability Lab

non-profit
https://aisilab.github.io/
aisilab
Activity Feed

AI & ML interests

Interpretability-informed control

Recent Activity

lgalke  authored a paper 2 days ago
Selecting The Most Informative Tokens in Natural Language Autoencoders
lgalke  authored a paper 2 days ago
Write Once, Run Everywhere: The Axon DSL for Shape-Safe and Framework-Agnostic LLM Architectures
lgalke  authored a paper 2 days ago
DFM Mimir v1: An Open HRM Delivering Frontier Performance at 1B Parameters Using Only Permissible Post-Training Data
View all activity

Papers

Selecting The Most Informative Tokens in Natural Language Autoencoders

Emergent Languages in Populations of Language Model Agents: From Token Efficiency to Oversight Evasion

View all Papers

Lukas Galke Poech's profile picture Stine Beltoft's profile picture William Brach's profile picture Federico Torrielli's profile picture Peter Schneider-Kamp's profile picture Gianluca Barmina's profile picture Filippo Tonini's profile picture

aisilab 's collections 1

Moltbook Models
  • filter-with-espresso/Qwen2.5-14B-Instruct-reddit-baseline-v3-high

    Updated Mar 17
  • filter-with-espresso/Qwen2.5-14B-Instruct-reddit-baseline-v2-low

    Updated Mar 16
  • filter-with-espresso/Qwen2.5-14B-Instruct-reddit-baseline-v1

    Updated Mar 16
  • filter-with-espresso/Qwen2.5-14B-Instruct-moltbook-finetune-v9

    Updated Mar 15 • 2
Moltbook Models
  • filter-with-espresso/Qwen2.5-14B-Instruct-reddit-baseline-v3-high

    Updated Mar 17
  • filter-with-espresso/Qwen2.5-14B-Instruct-reddit-baseline-v2-low

    Updated Mar 16
  • filter-with-espresso/Qwen2.5-14B-Instruct-reddit-baseline-v1

    Updated Mar 16
  • filter-with-espresso/Qwen2.5-14B-Instruct-moltbook-finetune-v9

    Updated Mar 15 • 2
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs