[devnexus 2026] it’s up to java developers to fix enterprise ai

Speaker: Rod Johnson (@springrod)

See the DevNexus live blog table of contents for more posts


General

  • Personal assistant approaches don’t work in the enterprise
  • Hype can be distracting
  • AI conversation driven by people not interested in/don’t understand the enterprise
  • Some things work; some don’t.
  • Change is fast. ex: Clause over last 6 months

Personal assistants

  • Personal assistants use cases work best
  • Coding assistants are a type of personal assistant
  • Valuable because you are at computer and say yes/no
  • This doesn’t walk in enterprise. Business process/chatbot with public can’t do take backs
  • Broad, flexible, human in the loop, tolerance for error, chat oriented. (By contrast business processes need to be specific, predictable, automated, reliable, workflow oriented)

Claude Code Execution Process

  • Analyze Request
  • Create to do list
  • Work through tasks
  • Test each step
  • Ensure integration

Unavoidable Challenges

  • Non deterministic
  • Hallucinations
  • Prompt engineering is alchemy. Throw things in vs engineering
  • Slow and expensive to run at scale
  • Difficult to test and validate

Avoidable Challenges

  • Top down mandates
  • “AI all the things”/AI for the sake of AI – should be doing incrementally
  • Wrong people controlling AI strategy. Data science group doesn’t always understand the business
  • Greenfield fallacy – business systems/workflows already exist. Domain context exists.
  • “This time is different” – no matter how shiny a new technology is; doesn’t change everything.

Instructive Open Claw Problems

  • Lack of structure – relies on markdown.
  • Token bloat and very high cost
  • Needs to compress context frequently which can change meaning/introduce risk of errors
  • Unpredictable especially as context grows
  • Lack of explainability
  • Exposed infrastructure risk article – egads. This is scary!

How to succeed

  • Attack non determinism – make as predictable as possible by breaking complex tasks into small steps; smaller prompts, less tools in context, mix code in for some steps, create guardrails, (Also saves money because some steps can use a cheaper LLM)
  • Integrate with what works – connect to existing system, leverage current domain expertise/coding skills, build on proven infrastructure, incremental
  • Build structure to LLM Interactions – don’t talk English if can avoid; include as much structure as possible. Ask for format of structured data.

Testing

  • Unit testing can find that you sent the wrong prompt (or implemented wrong)
  • Integration testing – test with real LLM but fake data – ex: test containers

Domain Integrated Context Engineering (DICE)

  • Context engineering more broad than prompt engineering
  • Bridges LLM/business system
  • Helps structure input and output
  • Domain objects
  • Integrate with existing domain models
  • Structure is a continuum from Open Claw (autonomous/unstructured) to old fashioned code. In between is Claude, MCP, agent frameworks and deterministic planning.
  • Embabel is the agent framework/deterministic planning level

What do as Java developers

  • Gen AI works best alongside existing systems
  • Your data/domain models/business rules
  • AI should extend your capabilities not replace them.
  • Think integration, not greenfield
  • Java skills undervalued to this point
  • Every Java developer should know both Java and Python [I do; yay]

Python vs Java

  • Don’t just imitate Python approaches
  • Build better – look at prior art (Python), leverage domain experience, apply architecture experience, bring strengths to Gen AI, create better frameworks, lead
  • Python – great for data science (data science != gen ai), scripting, prototyping
  • JVM – excels at enterprise grade applications

Embabel

  • Directly addresses key Gen AI failure points
  • Key innovation is deterministic planning [Python frameworks do not do this]
  • Goal Oriented Action Planning (GOAP)
  • Predictable/explainable execution
  • Actions and goals create extensible system
  • Includes a server; knows what up to.
  • Knows about all deployed capabilities and can extend
  • Builds up understanding of domain
  • Will become AI fabric of enterprise
  • Framework written in Kotlin; put a lot of effort into making sure easy to use from Java.
  • Most examples in Java and most of users/community are Java
  • Builds on existing stack.

Unfolding Tools

  • While better to have samller steps with less tools, sometimes you need a lot of tools
  • Tools use a lot of context and can confuse the LLM
  • Unfolding saves tokens and improves accuracy
  • Exposes a single top level tools. When invoked it expands to show children. Like Russian nesting dolls
  • Works by rewriting message history within agentic loop

Agentic Tools

  • Like supervisor pattern in Python framework, but more deterministic
  • Eposes single top level tool that coordinates lower level tools
  • Advanced implementations allow controlling order

RAG

  • Currently pipeline RAG. Do query , no feedback, hard to adjust
  • Future is agentic RAG – context aware multi step search with self-correction. LLM has more autonomy. Can do more searches: text, vector, expand chucks, etc

Rod wrote blog post: You can build better AI Agents in Java than Python

My take

After hearing about one shotting and exaggerations on social media, having a more balanced take was great. I especially appreciated the *whys*. I also liked the “what can you do” to use AI more safety problem.

[devnexus 2026] live blog table of contents

See my live blogs from the event

Thursday

Friday

[kcdc 2025] designing for behavioral change – the science behind habit-forming products

Speaker: Preston Chandler

For more see the table of contents


General

  • Why do some products become second nature while others are forgotten – valuable, fun, etc

Habit Loop

  • Cue -> Response -> Reward
  • If you put a golf ball near a nest, a goose will pull it into nest. Maximizes number of chicks from when egg rolls out of nest

Hook Model

  • Trigger -> Action -> Variable reward -> investment
  • Investment can be effort/time/money
  • Consultants expensive. If free, wouldn’t care about. “That was just $100 of advice”
  • Variable rewards are more appealing than predictable ones. ex: gambling
  • Some things need to be predictable – ex: excel formula

B=MAP

  • behavior = motivation * ability * prompt
  • Cathedral in Milan – had to sign up for entry with a QR code. Prompt was QR code. Motivated to get in. Couldn’t get website to work after 20-30 minutes

Effort vs Reward

  • Amazon – easy – buy now button, reward by getting stuff faster, microtransactions, made easy for you to give them money.
  • Tiktok – easy – just scroll down and get gratification. Variable reward; not every video good. Also dark pattern.
  • US Treasury – hard. Keyboard where click each letter and not in order. Changed since
  • hard website – abandon
  • AT&T – expensive. Negative reward compared to others. 8 hours to leave service. Multiple calls to customer service. People will never go back if left dissatisfied
  • Rewards – money, time, scrolling motivation
  • Checklists motivate most people, satisfaction of moving as done
  • Line of sight goals are motivating. Ex: daily goals, gold coins
  • Different people motivated by different things

Dark Patterns

  • Sign up for newsletter and get 30% in
  • Confusing radio buttons on whether to opt in
  • Link with very little contrast to background so can barely see
  • Company and user incentives not aligned

Voice Assistant

  • Use for music, timer, shopping
  • Sticky because personalizable to you

DuoLingo

  • Motivated to keep streak alive. Child said didn’t have enough time to finish homework. Said ok because went zoo. But wanted to keep streak
  • Easy to pick up, don’t need a lot of time
  • Bird will look angry and shame you – dark pattern
  • Constantly upselling – dark pattern

Exercise

  • For trigger clarity, action simplicity, reward value and investment payoff, think about obstacle today and how make better

Other

  • Behavior is deisgnable – ex: clear trigger, low effort
  • Ethics = engagement + trust
  • Small changes can have a big impact. If hose squished, have a constraint and hardly any water goes though. Must fix that to improve

Playbook

  • Identify internal/external triggers
  • Minimize friction, simplify first action
  • Offer variable rewards tied to meaning
  • Encourage invementment, effort builds attachment
  • Align outcomes with user values

Creativity

  • Chore Kanban
  • Have ChatGPT make budget a Shakespearean sonnet

My take

Great examples to understand ideas. Fun examples