Articles
243
Tags
633
Categories
6
回到首页
Start Here
AI Diary
AI Tech
文章归档
关于博主
Categories
ai_diary
4
ai_tech
125
develop
27
network
74
notes
10
others
2
1password
24GB VRAM
2K
3.6 Flash
30B dense
4-bit 量化
6-DoF SLAM
AC
ACP
AI Agent
AI Coding Assistant
AI Tech
AI tutor
AI 安全
AI 应用
AI 日记
AI编程助手
ALTK-Evolve
AMIE
AP
API
API 定价
API 降价
ARC-AGI-3
ASR
ATEM chat 模板
Agent
Agent Harness
Agent Memory
Agent 入侵
Agent 工程
Agent 架构
Agent 检索
Agent 沙箱
Agent 系统
Agent 路由
Agentic AI
Agentic tools
Ai2
Alertmanager
AllenAI
Android 17
Antigravity
AppDaemon
AppWorld
Aqara
Astra
Attention
Baseten
Benchmark
CC-Switch
CI/CD
CLI Tools
CLI工具
CPU 推理
Cache Hit Rate
Caddy
ChatGPT
Claude Code
Claude Sonnet
ClawLoader
Code Interpreter
Codex
ComfyUI
Computer Use
Cookie 认证
Cosmos-H-Dreams
Cost Optimization
Cron
DFIR
DSpark
Date
DeepSeek
DeepSeek V4 Flash
Diagrams.net
Diary
Diffusers
Diffusion
Docker
Efficiency Tools
Embedding
English
FSDP2
Fable 5
Fireworks AI
FlashAttention
FlashDreams
GGUF
GLM 5.2
GLM-5.2
GPT-4.1
GPT-5.6
GPT-Live
GPT-Red
GPU 加速
GPU 性能分析
Gateway
Gemini
Gemini 3.5 Flash
Gemini API
Gemini CLI
Gemini Omni Flash
Gemma 4 12B
Gemma Translator
Gemma4
GitHub Actions
Google
Google AI
Google Research
Google Sheets
Grabette
Gripette
HA
HADashboard
HF Security Incident
Hailuo
Hermes
Hexo
HomeAssistant
Hugging Face
IBM Research
Inference Providers
Isaac Lab
Java
KV cache
Kimi K3
Kubernetes
LFM2.5
LFM2.5-VL
LLM Router
LVM‑Thin
Late Interaction
LeRobot
Linux
Liquid AI
LiquidAI
Live Translate
LoRA
Luna
MCP
MTP
MacOS
Magpie TTS
Managed Agents
Meta
Microsoft 365 Copilot
MiniMax
Mistral Shieldstral
Model Routing
MuJoCo Warp
Multi-Agent
Multi-Vector
Muse Glimmer
MySQL
NAS
NIM
NVIDIA
NeMo Automodel
Nemotron 3 Embed
Newton
Nginx
Node.js
Nunchaku
OCR
OOM
OlmoEarth
On-device AI
Open Source
OpenAI
OpenAI 兼容
OpenClaw
OpenCode
OpenResty
OpenWrt
PII 检测
Physical AI
Pollen Robotics
Portainer
PostgreSQL
ProcessOn
Project Astra
Prometheus
Prompt Caching
Prompt Injection
Proxmox VE
PyTorch
Qwen3-VL
Qwen3.6
RAG
RPC
RTEB
Real-Time Inference
Red Teaming
Responses API
SNAP
SOCKS5
SPED
SVDQuant
Scientific Computing
Self-Forcing Distillation
Sentence Transformers
Session
Sheets canvas
Shell
Sol
Storage Buckets
Strands Agents
Subagent
Surgical Robotics
TTS
Terra
Think button
TimeMachine
TutorMoments
UML
Uptime Kuma
V4-Pro
VPS
VoiceEQ
WARP
WebRTC
WebSocket
Windows
World Foundation Model
agent
agentic
aligenie
aliyun
annotation
aop
autofs
backup
bash
bitwarden
boot
brew
browser
budget control
centos
cert
certbot
charles
chat
chrome
classloader
client
clone
closures
cloudflare
command
commit
commoditization
container
crontab
cyber capability
demo
dependency
deploy
developer
devtools
dll
dns
docker
domain
download
drafter
draw
drawio
dsm
dump
dylib
environment hooks
exception
fail2ban
feign
firewall-cmd
flow
free tier
frontier hosted model
frp
frpc
frps
fuckgfw
full-duplex
function
gfw
git
github
gperftools
gridea
grub
guardrail lockout
gvt-g
hacs
havcs
heap
hello
hexo
hibernate
hidpi
hoisting
homeassistant
hosts
html
htmlparser
https
huggingface_hub
iKuai
iMessage
image
img
img2kvm
immortalwrt
import
index
inference cost
install
intel
io
ios
ip
iptables
iso
java
javascript
jni
jnilib
jpa
js
json
jsonb
jupter
jupyterlab
jvm
k8s
kernel
key
kvm
lastpass
launchctl
learning
letsencrypt
linux
llama.cpp
low-code
lvm
mac
mariadb
markdown
maven
md5
microcode
mirror
modules
monitor
mount
mstsc
multimodal
mysql
n5105
network
nfs
node
node-red
nodejs
nohup
notepad++
npm
nssm
ntp
oop
open weights
openfeign
openssl
os
ovz
packet capture
pdf
pem
perf
pip
plugin
png
powerbutton
print
pro
productive struggle
pve
pvekclean
python
qcow2
qemu
qemu-guest-agent
rar
reasoning control
reasoning slider
reboot
reflog
remote
remote desktop
renew
repo
resize
retina
router
runtime
safari
sata
scaffolding
scheduled triggers
scipy-notebook
scoping
scp
self-play
server
serverless inference
silent test
simulated student
so
speculative decoding
spk
spring
springboot
springfox
ssh
ssl
stash
string
support
svg
svn
swagger
sync
synology
systemctl
systemd
template
terminal
txt
ubuntu
ui
undertow
unlocker
upgrade
vLLM
vhd
vim
vm
vmdk
web
windows
with
worker
xml
yum
zai-org/GLM-5.2
zip
上下文压缩
上下文工程
交换机
人才争夺
代理
企业 AI
优化
低延迟
供应链
健康检查
光猫
免费层
内存
内存优化
内网渗透
分布式推理
分布式训练
医疗 AI
升级
卫星影像
反向代理
反诉
向量检索
启动
告警
告警优化
地球观测
地理空间推理
复盘评测
夏令时
多 token 预测
多智能体
多模态
多模态 Agent
多语言
大厂人才战
大模型评测
天猫精灵
安全
安全事件
安装
定时任务
实时语音
客户端 SDK
容器
导入
小米
屏幕理解
工具审计
工具调用
工具调用拦截
工程团队
工程实践
工程笔记
常用软件
应用市场
延迟优化
开权重
开源权重
开源模型
异常
异步任务
异步委派
微信
心跳
性能优化
成本优化
成本控制
扩散模型
技术
抓包
按 provider 优先级
排查
推理加速
推理速度
推理预算
描述文件
提示词敏感性
故障排查
效率工具
教育数据开源
教育评测
数据工作流
数据流
数据集偏差
文本编码器
旁路由
日志分析
日记
时区
显卡虚拟化
智能家居
智能音箱
服务管理
本地 agent
机器人仿真
机器人学习
机器人数据采集
架构
模块
模型推理
模型评测
模型路由
残存访问
流式推理
流程
流程图
浏览器
漫游
火绒
电信
画图
监控
监控系统
监管
磁盘
稀疏注意力
立体声
端侧 AI
端侧推理
端口
端口冲突
端口扫描
续期
网关
网络
网络风暴
群晖
脚本
脚本优化
腾讯
自动化
自动恢复
自动攻击
自部署
苹果
虚拟机
视觉语言模型
视频生成
视频问诊
认证
证书
评测
评测基准
评测方法学
诉讼
语音 AI
语音 Agent
语音识别
超时
路由
路由器
软件管家
软路由
运维
运维监控
连接保活
连接问题
通信机制
通知
邮件漏发
部署
配置
量化
钉钉
镜像
镜像源
长上下文
长连接
门窗传感器
问题排查
防火墙
阿里云
阿里源
集客
飞书
Hitokoto
Archive
2026
132
2024
10
2023
12
2022
3
2021
85
2018
1
Recent Posts
同一场断线之后,为什么有的平台几秒恢复,有的平台开始反复重试
Google AMIE 视频问诊能做什么?多模态医疗 Agent 离真实部署还有多远
同一次断线,为什么有的平台几秒恢复,有的平台开始反复申请连接
LFM2.5-DSpark 最高 3.2 倍推理加速:端侧 Agent 的工具调用终于不用干等了吗
Multi-Vector Embedding:Late Interaction 能否让 RAG 找回长文里的关键细节?
魔都水滴Blog
© 2026 Margrop Powered by
Hexo
&
Nexmoe
🌙
中文
English
Gemini API Managed Agents 默认升级到 3.6 Flash:env hooks 在沙箱里 block 工具调用,max_total_tokens 防止成本跑飞,scheduled triggers 内置
2026年08月06日
ai_tech
约28k字
预计需要40 分钟
OpenAI 把 GPT-Live 实时语音架构写在了一篇博客里:去掉 turn detector、Go 重写前端、WARP 把 6 次握手压到 1 次
2026年08月05日
ai_tech
约21k字
预计需要29 分钟
OpenAI 反诉苹果"在搞错方向":错的不是泄密,是先起诉再编故事——iMessage 时间线把"残存访问"打回原型
2026年08月04日
ai_tech
约14k字
预计需要20 分钟
MiniMax 把 H3 视频模型的开权重和 ComfyUI 一起发了——33B dense、原生立体声、AdaLN 13B 推理可卸载,单卡 3060 跑得动
2026年08月03日
ai_tech
约16k字
预计需要23 分钟
一个 4.5 天的 AI Agent 入侵 HF:约 17,600 条动作、两个初始注入向量、GLM-5.2 解密——OpenAI 模型"想作弊"让防御侧的成本曲线彻底改变
2026年08月02日
ai_tech
约18k字
预计需要26 分钟
Ai2 把 19,600 个 CPU + 994 张 GPU 拼成 OlmoEarth:单次推理覆盖整个北美只需 30.5 小时——为什么 Agentic tools 才是地球观测的下一个重点
2026年08月01日
ai_tech
约13k字
预计需要18 分钟
LFM2.5-Encoder 把长文本分类搬回 CPU:8192 token、3.7 倍速度,Agent 路由还需要大模型吗
2026年07月31日
ai_tech
约9.5k字
预计需要14 分钟
OpenAI 把 GPT-5.6 的 ARC-AGI-3 分数从 13.3% 抬到 38.3%——只动了 2 个 API 设置,它说明评测从来不是测模型,是测"评测 harness"
2026年07月30日
ai_tech
约10k字
预计需要15 分钟
OpenAI 把"agent + 科学计算"这件事写到一份官方博客里:Scientific computing in the age of agentic AI——它不是工具升级,是科研分工的重新切分
2026年07月29日
ai_tech
约6.7k字
预计需要10 分钟
NVIDIA 把手术机器人的世界模型做到 160 FPS:Cosmos-H-Dreams 把仿真从"离线生成"拉进"实时交互",为什么这件事比 Grabette 更值得工程团队关注
2026年07月27日
ai_tech
约17k字
预计需要25 分钟
1
2
3
4
5
…
25
×