Found 393 bookmarks
Newest
(21) Kun Chen on X: "since it’s been a good day for anthropic with a strong opus 5.5 release, i’m going to highlight one more thing that may not be obvious across xai, openai and anthropic - 1. anthropic is the only provider that does not charge 2x for long context requests even at 1M context, https://t.co/fUB97iR0y6" / X
(21) Kun Chen on X: "since it’s been a good day for anthropic with a strong opus 5.5 release, i’m going to highlight one more thing that may not be obvious across xai, openai and anthropic - 1. anthropic is the only provider that does not charge 2x for long context requests even at 1M context, https://t.co/fUB97iR0y6" / X
·x.com·
(21) Kun Chen on X: "since it’s been a good day for anthropic with a strong opus 5.5 release, i’m going to highlight one more thing that may not be obvious across xai, openai and anthropic - 1. anthropic is the only provider that does not charge 2x for long context requests even at 1M context, https://t.co/fUB97iR0y6" / X
(21) Lance Martin on X: "useful tip for Opus 5.5: run “/claude-api prompt-audit” in Claude Code. this checks you skills, agent.md, Claude.md, prompts and removes anti-patterns that hobble frontier models. i updated the skill w/ the latest Opus 5.5 guidance. https://t.co/sBCIwcOCv1" / X
(21) Lance Martin on X: "useful tip for Opus 5.5: run “/claude-api prompt-audit” in Claude Code. this checks you skills, agent.md, Claude.md, prompts and removes anti-patterns that hobble frontier models. i updated the skill w/ the latest Opus 5.5 guidance. https://t.co/sBCIwcOCv1" / X
·x.com·
(21) Lance Martin on X: "useful tip for Opus 5.5: run “/claude-api prompt-audit” in Claude Code. this checks you skills, agent.md, Claude.md, prompts and removes anti-patterns that hobble frontier models. i updated the skill w/ the latest Opus 5.5 guidance. https://t.co/sBCIwcOCv1" / X
(7) Thariq on X: "we’re thinking of killing plan mode and using the shift+tab hotkey to adjust effort levels I don’t think the models need plan mode anymore, but if you’re a plan mode diehard would love to get your feedback on why" / X
(7) Thariq on X: "we’re thinking of killing plan mode and using the shift+tab hotkey to adjust effort levels I don’t think the models need plan mode anymore, but if you’re a plan mode diehard would love to get your feedback on why" / X
·x.com·
(7) Thariq on X: "we’re thinking of killing plan mode and using the shift+tab hotkey to adjust effort levels I don’t think the models need plan mode anymore, but if you’re a plan mode diehard would love to get your feedback on why" / X
(21) Stencil on X: "omp 18.2 is now out! - Much speedier startup, full TUI booting faster than Codex, CC, and Pi, - A new indicator, detecting and warning against gateways claiming mismatching frontier models, - Ollama web search, refreshed /skill UX, tiered OpenRouter routing & much more https://t.co/srtaoMCTnf" / X
(21) Stencil on X: "omp 18.2 is now out! - Much speedier startup, full TUI booting faster than Codex, CC, and Pi, - A new indicator, detecting and warning against gateways claiming mismatching frontier models, - Ollama web search, refreshed /skill UX, tiered OpenRouter routing & much more https://t.co/srtaoMCTnf" / X
·x.com·
(21) Stencil on X: "omp 18.2 is now out! - Much speedier startup, full TUI booting faster than Codex, CC, and Pi, - A new indicator, detecting and warning against gateways claiming mismatching frontier models, - Ollama web search, refreshed /skill UX, tiered OpenRouter routing & much more https://t.co/srtaoMCTnf" / X
(21) Hassan on X: "Introducing Inspo. A design MCP server for Claude Code, Codex, and OpenCode. It searches 800+ beautiful websites and finds relevant design inspiration for your coding agent. Install → npx inspo-mcp install https://t.co/iyw777zrgj" / X
(21) Hassan on X: "Introducing Inspo. A design MCP server for Claude Code, Codex, and OpenCode. It searches 800+ beautiful websites and finds relevant design inspiration for your coding agent. Install → npx inspo-mcp install https://t.co/iyw777zrgj" / X
·x.com·
(21) Hassan on X: "Introducing Inspo. A design MCP server for Claude Code, Codex, and OpenCode. It searches 800+ beautiful websites and finds relevant design inspiration for your coding agent. Install → npx inspo-mcp install https://t.co/iyw777zrgj" / X
alibaba/open-code-review: Fast, efficient, battle-tested at Alibaba's scale. Hybrid architecture code review tool: deterministic pipelines + LLM Agent, precise line-level comments, built-in multi-language ruleset (NPE, thread-safety, XSS, SQL injection), OpenAI & Anthropic compatible.
alibaba/open-code-review: Fast, efficient, battle-tested at Alibaba's scale. Hybrid architecture code review tool: deterministic pipelines + LLM Agent, precise line-level comments, built-in multi-language ruleset (NPE, thread-safety, XSS, SQL injection), OpenAI & Anthropic compatible.
·github.com·
alibaba/open-code-review: Fast, efficient, battle-tested at Alibaba's scale. Hybrid architecture code review tool: deterministic pipelines + LLM Agent, precise line-level comments, built-in multi-language ruleset (NPE, thread-safety, XSS, SQL injection), OpenAI & Anthropic compatible.
(21) ClaudeDevs on X: "New in Claude Code: claude plugin eval See what value your plugin is adding, or if it needs more work. You can create test cases, run your plugin or skill against those test cases, score those runs, then run each case again without the plugin to see the differences. https://t.co/qfPU6WHueV" / X
(21) ClaudeDevs on X: "New in Claude Code: claude plugin eval See what value your plugin is adding, or if it needs more work. You can create test cases, run your plugin or skill against those test cases, score those runs, then run each case again without the plugin to see the differences. https://t.co/qfPU6WHueV" / X
·x.com·
(21) ClaudeDevs on X: "New in Claude Code: claude plugin eval See what value your plugin is adding, or if it needs more work. You can create test cases, run your plugin or skill against those test cases, score those runs, then run each case again without the plugin to see the differences. https://t.co/qfPU6WHueV" / X
(12) Anthropic on X: "We're publishing our most detailed threat intelligence report to date. It covers how people tried to misuse Claude—for cyberattacks, influence operations, surveillance, biology, and building weapons—and how we found and stopped them. We disrupted every operation in the report," / X
(12) Anthropic on X: "We're publishing our most detailed threat intelligence report to date. It covers how people tried to misuse Claude—for cyberattacks, influence operations, surveillance, biology, and building weapons—and how we found and stopped them. We disrupted every operation in the report," / X
·x.com·
(12) Anthropic on X: "We're publishing our most detailed threat intelligence report to date. It covers how people tried to misuse Claude—for cyberattacks, influence operations, surveillance, biology, and building weapons—and how we found and stopped them. We disrupted every operation in the report," / X
(11) Lee Robinson on X: "We just rolled out CursorBench 4.0! It includes new tasks for how well models follow instructions, work on challenging projects over time, and is more difficult than before (so all models score lower). https://t.co/cYVgRBEFWE" / X
(11) Lee Robinson on X: "We just rolled out CursorBench 4.0! It includes new tasks for how well models follow instructions, work on challenging projects over time, and is more difficult than before (so all models score lower). https://t.co/cYVgRBEFWE" / X
·x.com·
(11) Lee Robinson on X: "We just rolled out CursorBench 4.0! It includes new tasks for how well models follow instructions, work on challenging projects over time, and is more difficult than before (so all models score lower). https://t.co/cYVgRBEFWE" / X
(21) ClaudeDevs on X: "You can pop out any pane in the Claude Code desktop app into its own window. Drag the diff or terminal to a second screen while Claude keeps working in the main window, then dock it back whenever you want. You can also run sessions side-by-side or stacked. https://t.co/5BF38zm4ky" / X
(21) ClaudeDevs on X: "You can pop out any pane in the Claude Code desktop app into its own window. Drag the diff or terminal to a second screen while Claude keeps working in the main window, then dock it back whenever you want. You can also run sessions side-by-side or stacked. https://t.co/5BF38zm4ky" / X
·x.com·
(21) ClaudeDevs on X: "You can pop out any pane in the Claude Code desktop app into its own window. Drag the diff or terminal to a second screen while Claude keeps working in the main window, then dock it back whenever you want. You can also run sessions side-by-side or stacked. https://t.co/5BF38zm4ky" / X
(8) Tibo on X: "Demand for Astra is really unprecedented. We're pulling all the levers possible to sustain the demand, but I've not seen anything like it until now and we went through very steep growth before. Priority will always be to keep excellent service for existing users, but we might" / X
(8) Tibo on X: "Demand for Astra is really unprecedented. We're pulling all the levers possible to sustain the demand, but I've not seen anything like it until now and we went through very steep growth before. Priority will always be to keep excellent service for existing users, but we might" / X
·x.com·
(8) Tibo on X: "Demand for Astra is really unprecedented. We're pulling all the levers possible to sustain the demand, but I've not seen anything like it until now and we went through very steep growth before. Priority will always be to keep excellent service for existing users, but we might" / X
(20) Tibo on X: "To calibrate you all on which reasoning effort to use for Astra, know that GPT-6 Astra on low performs better than GPT-5.6 Sol on high. If you were using high reasoning efforts with Sol and were happy, I suggest you move down to low or medium for Astra." / X
(20) Tibo on X: "To calibrate you all on which reasoning effort to use for Astra, know that GPT-6 Astra on low performs better than GPT-5.6 Sol on high. If you were using high reasoning efforts with Sol and were happy, I suggest you move down to low or medium for Astra." / X
·x.com·
(20) Tibo on X: "To calibrate you all on which reasoning effort to use for Astra, know that GPT-6 Astra on low performs better than GPT-5.6 Sol on high. If you were using high reasoning efforts with Sol and were happy, I suggest you move down to low or medium for Astra." / X
(21) Guanlan Dai on X: "A year ago the question was which model. Now it's which harness. Pi, Exo, Claude Code, Codex, DeepSeek Harness and 4 others. Same model, same tasks, same runtime. 360 runs, 2 billion tokens. Pass rates: 50% to 67%. Cost per pass: $1.05 to $18.34. Introducing FrontierHarness https://t.co/4QCzPyCmYc" / X
(21) Guanlan Dai on X: "A year ago the question was which model. Now it's which harness. Pi, Exo, Claude Code, Codex, DeepSeek Harness and 4 others. Same model, same tasks, same runtime. 360 runs, 2 billion tokens. Pass rates: 50% to 67%. Cost per pass: $1.05 to $18.34. Introducing FrontierHarness https://t.co/4QCzPyCmYc" / X
·x.com·
(21) Guanlan Dai on X: "A year ago the question was which model. Now it's which harness. Pi, Exo, Claude Code, Codex, DeepSeek Harness and 4 others. Same model, same tasks, same runtime. 360 runs, 2 billion tokens. Pass rates: 50% to 67%. Cost per pass: $1.05 to $18.34. Introducing FrontierHarness https://t.co/4QCzPyCmYc" / X
(21) Z.ai on X: "GLM-5.3 is now open-weight. Our most capable model for agentic coding and cyber defense is now available to download, run, and customize. Weights: https://t.co/v1IbWMXxg4 Tech blog: https://t.co/ekQkO83jCv https://t.co/f8XlJksKyf" / X
(21) Z.ai on X: "GLM-5.3 is now open-weight. Our most capable model for agentic coding and cyber defense is now available to download, run, and customize. Weights: https://t.co/v1IbWMXxg4 Tech blog: https://t.co/ekQkO83jCv https://t.co/f8XlJksKyf" / X
·x.com·
(21) Z.ai on X: "GLM-5.3 is now open-weight. Our most capable model for agentic coding and cyber defense is now available to download, run, and customize. Weights: https://t.co/v1IbWMXxg4 Tech blog: https://t.co/ekQkO83jCv https://t.co/f8XlJksKyf" / X