Howardism · Vol. 03Plate I · No. 01
Writing, in order.
Pieces540Sections4Oldest10 Apr 2026Newest5 Oct 2026
Every article in the wiki, grouped by kind: Concept notes, Entity profiles, and Essay pieces. Hover any title for a preview; click to enter.
Plate · 01 of 04
368
concept notes
Concept, in order.
Agent Systems53
- 01· 10′
- 02· 16′
- 03· 13′
- 04Agent Documentation BehaviorAgent EngineeringDocumentationTelemetry+1· 20′
- 05· 30′
Evals & Benchmarks36
- 01· 13′
- 02The Price of Fixed CapabilityInference CostBenchmarksEvaluation Methodology+2· 15′
- 03Verbalized-Confidence Soft Scoring for LLM JudgesLLM As A JudgeCalibrationVerbalized Confidence+2· 16′
- 04· 18′
- 05· 20′
Agent Security33
- 01· 16′
- 02· 18′
- 03· 17′
- 04AI-Enabled Influence OperationsThreat LandscapeInfluence OperationsDeception+2· 15′
- 05AI-Enabled State SurveillanceThreat LandscapeSurveillanceGovernance+2· 15′
AI Economics & Labor32
- 01· 22′
- 02· 14′
- 03AI-Moderated Interviews: Adaptive Probing, Human Rapport, and Digital TwinsMarket ResearchHuman AI SubstitutionLLM Simulation+2· 13′
- 04· 19′
- 05· 25′
Model Capability & Training29
- 01Diversity Calibration Under SFTSFTDiversityMode Collapse+2· 7′
- 02· 14′
- 03· 16′
- 04· 15′
- 05· 18′
Superintelligence Trajectory29
- 01Silent Revision RateGovernanceAI PolicyTransparency+2· 14′
- 02· 32′
- 03· 19′
- 04· 28′
- 05· 34′
Product & Org20
- 01Build Instead of Buy Under Agentic CodingProcurementEnterprise DeploymentBuild Vs Buy+2· 14′
- 02Psychological Costs of AI AdoptionOrganizationAdoptionDeveloper Experience+3· 30′
- 03· 47′
- 04Prototype Fidelity After Cheap PolishDesign ProcessPrototypingResearch Agenda+1· 10′
- 05· 25′
Startup & Founder17
- 01· 18′
- 02· 24′
- 03· 8′
- 04· 37′
- 05· 46′
Interpretability13
- 01· 13′
- 02Interference WeightsInterpretabilitySuperpositionMechanistic Interpretability+1· 19′
- 03· 24′
- 04· 17′
- 05· 14′
Interaction & Multimodal11
- 01· 27′
- 02· 29′
- 03· 21′
- 04Why AI Lags at DesignDesignModel CapabilityVerifiability+1· 8′
- 05· 25′
Plate · 02 of 04
96
entity notes
Entity, in order.
- E01· 7′
- E02· 7′
- E03· 5′
- E04· 5′
- E05· 5′
- E06AccentureEntityOrganizationConsulting+1· 6′
- E07· 6′
- E08· 15′
- E09· 8′
- E10· 8′
Plate · 03 of 04
60
essay pieces
Essay, in order.
- S01· 11′
- S02· 14′
- S03· 9′
- S04· 15′
- S05· 18′
- S06· 11′
- S07· 18′
- S08· 16′
Plate · 04 of 04
16
index notes
Index, in order.
- I01· 21′
- I02· 127′
- I03· 22′
- I04· 6′
- I05· 24′
- I06· 36′
- I07· 30′
- I08· 21′
Plate · #
486
subjects
By subject, most cited.
Agent EngineeringEntityEmpiricalAnthropicGovernanceAlignmentEvaluationAI Coding WorkflowLLM ArchitectureSecurityDerivedMeasurementSafetyBenchmarksWorkforcePersonHarnessHuman AI CollaborationEvaluation MethodologyEconomicsTest Time ComputeCode ReviewInterpretabilityCapability EvaluationPost TrainingStartupLLM As A JudgeMulti AgentAI For MathematicsContext ManagementLLM ModelMonitoringOpenaiOrgEngineering MetricsMultimodalVerificationAI Native OrgOpen WeightsProduct ManagementRecursive Self ImprovementAgent OrchestrationCapability TrajectoryCode QualityFailure ModesFounderPlanningPrompt InjectionReinforcement LearningAI PolicyFormal MethodsKnowledge ManagementClaudeConstruct ValidityCostThreatsGoogleOrganizationPromptingReward HackingTheoryThreat LandscapeAI CodingAI RdCalibrationFrontier LabsIncident ResponseOpen SourceOrg DesignReasoningAgentsAI EvaluationAI NativeClaude CodeCoding AgentsForecastingGovernance WorkforceIdentityLLM EvaluationMCPPsychometricsReference MonitorReliabilityResearcherReward DesignSupply ChainTrainingType/entityZero TrustAccountabilityAgentic RlAuthorizationChain Of ThoughtCoordinationCybersecurityEvalsInference ScalingMemoryRetrievalSafeguardsSearchSkillsSynthetic DataTechnical DebtTechnology DiffusionTool UseTrust BoundaryAdaptive EvaluationAI LabArchitectureAutomationCode GenerationCompanyDeceptionEnterprise DeploymentHonestyMethodologyMoatsModel WelfareOptimizationOrchestrationOversightPractitioner OpinionPrdProduct OrgProduct StrategyRole EvolutionStartupsTasteTeam DesignToolTraining GamingWorkflow DesignAgent DeploymentAgent HarnessAgent RuntimeAgent SecurityAI AdoptionAsiAuthenticationCase StudyCodingCognitive LoadCompetitive StrategyDeep ResearchDeploymentDesignDeveloper WorkflowDual UseEducationExpertiseGo To MarketHRIntegrationLaborLeast PrivilegeMacroMisalignmentPersonaProductProduct ProcessPropagationPrototypingProvenanceResearch MethodologyRspScalingSelf ReportSoftware EconomicsStanfordSuperintelligenceTestingTraining EfficiencyAccess ControlAdoptionAI ControlAI CoworkingAI Native OrganizationBelief ModificationBenchmark IntegrityCareerCatastrophic RiskClassifiersCode AuthorshipCodexCognitionCompetitive LandscapeComputeCoworkCredentialsData ContaminationData SourceData WallDeepmindDefense In DepthDefensibilityDiscoveryDocumentDocument ParsingEngineering LeadershipEquityGame TheoryGrowth DynamicsGtmHallucinationHuman CapitalIncentivesInferenceInterfaceInterpretability ResearcherIntrospectionLimitsLLM CharacterMental ModelMetrModelModel ImprovementModel RoutingModel SpecPermissionsPlatformPrivacyProbesProcessProduct ArchitectureProduct TasteProductionProductivityProtocolQualityQuantizationReasoning ModelsRed TeamingResearch ProgramRLHFScaling LawsScientific DiscoverySFTSoftware ArchitectureSoftware DesignStandardsSynthesisSystemTask SpecificationUniversal AIValue ModelVendorVendor ClaimVerifiabilityVulnerability ResearchY Combinator
+235 one-off subjects
AblationAbstractionAdaptive TestingAgencyAgent ArchitectureAgent FrameworksAgent SafetyAgent SkillsAgentic CodingAgentic EvaluationAgileAIAI EconomyAI EducationAI EthicsAI SafetyAI ToolsAixiAlignment ResearcherArms ControlArtifactsAsset PricingAuditingAuthorshipAutoformalizationAutonomyBenchmarkBenchmark ValidityBuild Vs BuyBusiness StrategyCapabilityCausal TheoryCharacterCLI AgentCoding AgentCollaborationCompetitionComplexity TheoryComputer UseConcept DiscoveryConfirmation BiasConfused DeputyConsciousnessConsultingContainmentContamination AdjacentContradictionControl PlaneCooperationCopyrightCourseCovert ChannelCreativityCredit AssignmentCultureData FlywheelData PipelineData ProvenanceDebate MapDebuggingDecision MakingDefenseDefinition Of DoneDefinitionsDelegationDelimiter InjectionDemocratizationDesign PrincipleDesign ProcessDesign SystemsDeterminismDeveloper EndpointDeveloper ExperienceDiffusionDigital IntelligenceDistillationDiversityDocumentationEconomic ImpactEconomic ModelingEconomistEfficiencyEmbodied AIEmergent BehaviorEmpirical SeEngineering IntelligenceEpistemicsError AnalysisExploit DevelopmentFailure ModeFaithfulnessFinanceFine TuningFormal VerificationFoundersFrontier LabGeneralistsGeneralizationGoodhartGoogle DeepmindGrounded TheoryGroundingGroup AgencyGrpoGuardrailsHardwareHciHierarchyHiringHistoryHouseholdHuman AI SubstitutionHuman In The LoopIncidentsInference CostInference EfficiencyInfluence OperationsInformation Flow ControlInfrastructureInput ValidationInstruction FollowingIntelligence MeasureInteraction MultimodalInternal ToolingInverse ScalingIsolationJailbreakKnowledge SharingLatent CapabilityLeaderboardsLearning GuideLeverageLifecycleLLM CapabilitiesLLM SimulationLong HorizonLoopsLow RankMarket DesignMarket ResearchMatrix CompletionMeasurement ValidityMechanistic InterpretabilityMediaMemorizationMessaging PlatformMetacognitionMethodMetricsMicrosoftMidtrainingMigrationMode CollapseModel CapabilityModel OrganismsModel PoisoningMultilingualMvpNational SecurityNetflixNous ResearchNvidiaOauthObservabilityOff HostOff Policy CorrectionOpen Ended DiscoveryOpinionsOrganizational TheoryOwaspPareto FrontierPersonal Knowledge BasePhilosophy Of SciencePlatform EngineeringPm SkillPolicyPrediction Powered InferencePrice TrendsPricingPrincipleProcurementProduction EvaluationPrompt EngineeringQualitative AnalysisRAGRelease DecisionsRepresentationsResearchResearch AgendaResearch AgentsResearch LabRevenue Per EmployeeReviewReward ModelingReward ModelsRl Post TrainingRole ShiftRubricsSafety EvaluationSandboxingScale StageShutdownSingularitySkills DevelopmentSoarSoftwareSoftware CraftSoftware DevelopmentSoftware ParadigmStatic AnalysisStatisticsSubagentsSuperpositionSurveillanceSystemsTalentTalent DensityTaxonomyTeam DynamicsTelemetryTest GenerationThird Party OversightThreat IntelligenceThreat ModelTool PoisoningTransparencyTyped DecisionsUncertaintyUnit EconomicsUnlearningValidationVenture CapitalVerbalized ConfidenceVersioningWorkflow